TurboQuant crushes Gemma 4 quant benchmarks
TurboQuant KV cache quantization excels on Gemma 4 26B in llama.cpp on M4 Pro, matching q4_0 quality at ~3.1 bits/K with 34% long-context speedup. Per-layer outlier-aware K quant outperforms public forks on Qwen PPL at lower bit rates.







