Search

Tag: #qwen24 results

ByteShape Qwen 3.5 9B Quants Guide

ByteShape Qwen 3.5 9B Quants Guide

ByteShape released quantized versions of Qwen 3.5 9B with benchmarks across GPUs like 5090/4080 and CPUs. Key GPU picks: 5.10 bpw baseline, 4.43 bpw balanced, 3.60 bpw fast. Blog offers interactive graphs for hardware-specific selection; first of more Qwen drops.

Reddit r/LocalLLaMACommunityMar 31#quantization#benchmarks#qwen
🤖

Uncensored Qwen3.5-4B Aggressive GGUF Drops

An uncensored version of the new Qwen3.5-4B model has been released in GGUF format with zero refusals out of 465 tests. It retains full capabilities, supports multimodal inputs, and offers various quants from 2.6GB to 7.9GB. Upcoming uncensored variants for larger Qwen3.5 sizes are in progress.

Reddit r/LocalLLaMACommunityMar 3#uncensored#gguf#multimodal
Qwen 3.5 Plus Launches on AI Gateway

Qwen 3.5 Plus Launches on AI Gateway

Qwen 3.5 Plus is now available on Vercel's AI Gateway, featuring a 1M context window and built-in adaptive tool use. It excels in agentic workflows, coding, web development, and multimodal tasks, outperforming Qwen 3 VL in scientific problem-solving and visual reasoning. Developers can access it via AI SDK by setting the model to alibaba/qwen3.5-plus.

Vercel NewsMediaFeb 16#launch#qwen#35-plus
Page 1 of 3