Address CodeRabbit review on #15410: clear quant_format/layout_type when a module reloads an unquantized weight (previously stale state made Linear.forward take the quantized path against a plain Parameter), and raise instead of silently loading raw bytes when an Embedding's inferred format is NVFP4, which the embedding path can't dequantize. |
||
|---|---|---|
| .. | ||
| test_mixed_precision.py | ||