| ▲ | danielmarkbruce 7 hours ago | |||||||
You are conflating post training quantization and low bit training. | ||||||||
| ▲ | kadushka 6 hours ago | parent [-] | |||||||
That's what I meant - we are currently use fp4 formats for training, and we cannot quite get away with that, despite dynamic quant and small block size - we still have to use quite a bit of higher precision (fp8 or even fp16) in various model components. | ||||||||
| ||||||||