Search

Tag: #inference127 results

Reka AI 舉辦 Edge 模型 AMA

Reka AI 舉辦 Edge 模型 AMA

Reka AI 在 r/LocalLLaMA 舉辦首場 AMA,由 Reka Edge 模型的研究領導參與。現場問答於 3 月 25 日上午 10 點至中午 12 點 PST,之後將非同步回覆問題。焦點在於實體世界應用、研究方向及最新模型。

Reddit r/LocalLLaMACommunityMar 24#ama#real-world-ai#inference
Nvidia Rubin Speeds MoE Inference

Nvidia Rubin Speeds MoE Inference

Nvidia's Rubin architecture features advanced NVLink interconnects for agentic AI and massive-scale MoE model inference at up to 10x lower cost per token. Combined with Groq's lightning-fast inference, it addresses latency crises in real-time AI reasoning. This shift enables frontier intelligence without brute-force compute scaling.

VentureBeatMediaFeb 15#launch#nvidia#rubin
OpenAI's Cerebras Instant Code Model

OpenAI's Cerebras Instant Code Model

OpenAI launched GPT-5.3-Codex-Spark on Cerebras chips for near-instant code generation at 1000+ tokens/sec. First major non-Nvidia inference partnership complements GPUs for low-latency. Available to Pro users via apps and extensions.

VentureBeatMediaFeb 12#launch#openai#codex-spark
Page 13 of 13