Minimax M2.5 GGUF Quants Disappoint
GGUF quantizations (Q4 to Q1) of Minimax M2.5 severely underperform the original model, unlike robust Qwen 3.5 quants. Evaluations on H200 took 10-20 hours per model due to gibberish generation. Key lesson: quantization robustness varies by model.


