Search

10 results on this page

🤖

Qwen KV Precision Shows Real Quality Gaps

A community test on an AMD R9700 with ROCm reports noticeable quality and long-context retention differences between FP16 and q8_0 KV cache for Qwen3.8-27B. FP16 reportedly produces more careful structured output and maintains performance beyond 120k tokens, challenging the assumption that both formats are equivalent.

Reddit r/LocalLLaMACommunity19h ago#kv-cache#long-context#quantization
Why Long-Form Video Resists Full AI

Why Long-Form Video Resists Full AI

Long-form audiences remain highly sensitive to visible AI flaws, making full-stack AI production risky for films and television. Industry leaders are instead adopting human-led workflows where AI supports visual effects, previsualization, cleanup, quality control, and other production bottlenecks.

Tencent Gray-Tests Flagship Hunyuan Hy4

Tencent Gray-Tests Flagship Hunyuan Hy4

Tencent's Hunyuan Hy4 has reportedly appeared in the model selection list of the Yuanbao app under an expert-level label and with tool-use capabilities. It is positioned above Hy3 and alongside DeepSeek, following Tencent's recent statement that a larger-parameter Hy4 would launch soon with improved performance and multimodal abilities.

Reddit r/LocalLLaMACommunity14h ago#model-testing#tool-use#multimodal
Page 1