New Bartowski Quants for Qwen3.5 27B

๐ก450 t/s on 5060 Ti makes Qwen3.5 27B viable for local coding/debugging
โก 30-Second TL;DR
What Changed
Bartowski released new Imatrix quants for Qwen3.5 27B
Why It Matters
Enhances local deployment of large models on consumer NVIDIA GPUs, enabling faster inference for coding tasks without cloud dependency.
What To Do Next
Download Bartowski's IQ2_M quant from Hugging Face and benchmark on your 40-series GPU.
Key Points
- โขBartowski released new Imatrix quants for Qwen3.5 27B
- โขIQ2_M achieves 450 t/s pp, 20 t/s tg on RTX 5060 Ti at 128k ctx
- โขSuperior debugging performance vs other Q2 variants
- โขGPU power usage peaks at 170-175W during inference
๐ง Deep Insight
Background and context from public sources โ not the original article. 7 sources cited.
๐ Enhanced Key Takeaways
- โขQwen3.5-27B is a native vision-language model supporting text, image, and video inputs with a 262,144-token context length[1][2][3].
- โขModel released on February 24, 2026, by Alibaba's Qwen team alongside Qwen3.5-122B-A10B and Qwen3.5-35B-A3B variants[5].
- โขHosted API pricing is $0.30 per 1M input tokens and $2.40 per 1M output tokens[2][3].
๐ ๏ธ Technical Deep Dive
- โขNative vision-language dense model with linear attention mechanism for fast response times and balanced inference speed/performance[1][2][3].
- โขOverall capabilities comparable to larger Qwen3.5-122B-A10B model[1][2][3].
- โขUses Qwen3 tokenizer; 27 billion parameters; modalities: text + image + video โ text output[2].
- โขBenchmark scores include GPQA 85.8%, HLE 22.2%, SciCode 39.5%, LCR 67.3%, IFBench 75.6%, Tau2 93.9%, TerminalBench Hard 32.6%[3].
๐ฎ Future ImplicationsAI analysis grounded in cited sources
โณ Timeline
๐ Sources (7)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
Weekly AI Recap
Read this week's curated digest of top AI events โ
๐Related Updates
AI-curated news aggregator. All content rights belong to original publishers.
Original source: Reddit r/LocalLLaMA โ
This is a summary, not the original. Read the source, or get the weekly briefing.
Weekly AI briefing
One email a week. Unsubscribe anytime.