QAD Q4_K_M?

#8
by coder543 - opened

Unsloth has shown with the Gemma 4 QATs how useful it is to have dynamic quantization even with QAT/QAD. I assume the output of your QAD training was a bf16 model, not directly a Q4_0 model, which was then quantized to a basic Q4_0? Maybe you could at least quantize that to Q4_K_M as well?

Just a thought.

Sign up or log in to comment