๐ฆReddit r/LocalLLaMAโขStalecollected in 4h
Arc Pro B70 Trails RTX 3090 in llama.cpp Benchmarks
๐กIntel Arc Pro B70 benchmarks reveal gaps and SYCL wins vs RTX 3090 for llama.cpp
โก 30-Second TL;DR
What Changed
Arc Pro B70 Vulkan averages -71.1% prompt processing vs RTX 3090
Why It Matters
Highlights Intel Arc as viable but slower alternative to Nvidia for local LLM inference, with SYCL offering gains. Encourages testing Intel hardware for cost-sensitive setups. May accelerate Arc driver optimizations for AI workloads.
What To Do Next
Run llama-bench on your Arc GPU with Vulkan and SYCL to compare inference speeds.
Who should care:Developers & AI Engineers
Key Points
- โขArc Pro B70 Vulkan averages -71.1% prompt processing vs RTX 3090
- โขSYCL on Arc boosts some models like Gemma-4-E2B by +50.3% over Vulkan
- โขToken generation: Arc SYCL excels in Qwen models, up +160% in one case
- โขOOM on Qwen3-Coder-30B for RTX 3090, Arc handles it
- โขTested on NixOS host and Ubuntu Docker for SYCL
๐ฐ
Weekly AI Recap
Read this week's curated digest of top AI events โ
๐Related Updates
AI-curated news aggregator. All content rights belong to original publishers.
Original source: Reddit r/LocalLLaMA โ