
Qwen3.8-Flash-Next Runs on 48GB Mac
A Reddit user reports running the 104GB Qwen3.8-Flash-Next model on a Mac with 48GB of memory. The setup reportedly achieves approximately 12 tokens per second, demonstrating practical local inference despite the model exceeding system memory.
Reddit r/LocalLLaMA · 14d ago





















