Search

Tag: #moe72 results

Qwen3.5-27B 領先家族基準測試

Qwen3.5-27B 領先家族基準測試

ArtificialAnalysis.ai 在智慧指數、程式設計指數與代理指數上,將 Qwen3.5-27B 評為高於 Qwen3.5-122B-A10B 與 Qwen3.5-35B-A3B。小型模型在所有類別勝過更大 MoE 兄弟模型。

Reddit r/LocalLLaMACommunityFeb 26#benchmarks#moe#qwen
Nvidia Rubin Speeds MoE Inference 10x Cheaper

Nvidia Rubin Speeds MoE Inference 10x Cheaper

Nvidia's Rubin platform features advanced NVLink interconnects to accelerate agentic AI, reasoning, and massive-scale MoE model inference at up to 10x lower cost per token. The article analogizes tech growth to a pyramid's limestone blocks, highlighting shifts from CPUs to GPUs and now efficient architectures. Groq complements this with ultra-fast inference to solve latency issues in real-time AI.

VentureBeatMediaFeb 15#launch#nvidia#rubin
JD Open-Sources 48B MoE JoyAI-Flash Model

JD Open-Sources 48B MoE JoyAI-Flash Model

JD.com open-sourced JoyAI-LLM-Flash, a 48B total parameter MoE model with 3B active params, pre-trained on 20T tokens. It excels in knowledge, reasoning, coding, and agents using FiberPO framework and Muon optimizer. Features 1.3x-1.7x throughput gains via dense MTP.

IT之家MediaFeb 15#launch#jd-com#joyai-llm-flash
Page 7 of 8