SourceStalecollected in 4h

Arc Pro B70 Trails RTX 3090 in llama.cpp Benchmarks

PostLinkedIn
🦙Read original on Reddit r/LocalLLaMA
#benchmarks#vulkan#sycl#ggufarc-pro-b70nvidiartx-3090intelarc-pro-b70llama.cpp

💡Intel Arc Pro B70 benchmarks reveal gaps and SYCL wins vs RTX 3090 for llama.cpp

⚡ 30-Second TL;DR

What Changed

Arc Pro B70 Vulkan averages -71.1% prompt processing vs RTX 3090

Why It Matters

Highlights Intel Arc as viable but slower alternative to Nvidia for local LLM inference, with SYCL offering gains. Encourages testing Intel hardware for cost-sensitive setups. May accelerate Arc driver optimizations for AI workloads.

What To Do Next

Run llama-bench on your Arc GPU with Vulkan and SYCL to compare inference speeds.

Who should care:Developers & AI Engineers

Key Points

  • Arc Pro B70 Vulkan averages -71.1% prompt processing vs RTX 3090
  • SYCL on Arc boosts some models like Gemma-4-E2B by +50.3% over Vulkan
  • Token generation: Arc SYCL excels in Qwen models, up +160% in one case
  • OOM on Qwen3-Coder-30B for RTX 3090, Arc handles it
  • Tested on NixOS host and Ubuntu Docker for SYCL
📰

Weekly AI Recap

Read this week's curated digest of top AI events →

👉Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: Reddit r/LocalLLaMA

This is a summary, not the original. Read the source, or get the weekly briefing.

The weekly digest

One email a week. Unsubscribe anytime.