
Opus 4.7 快速模式現於 AI Gateway
Claude Opus 4.7 的快速模式現已在 Vercel AI Gateway 的研究預覽中可用,提供約 2.5 倍更快的輸出 token 生成,保留完整模型智能。此實驗功能定價為標準 Opus 費率的 6 倍,輸入 $30/百萬 token,輸出 $150/百萬 token。可透過 provider options 或 Claude Code 環境變數啟用。
Tag: #fast-inference6 results

Claude Opus 4.7 的快速模式現已在 Vercel AI Gateway 的研究預覽中可用,提供約 2.5 倍更快的輸出 token 生成,保留完整模型智能。此實驗功能定價為標準 Opus 費率的 6 倍,輸入 $30/百萬 token,輸出 $150/百萬 token。可透過 provider options 或 Claude Code 環境變數啟用。

DeepMind 推出最新圖像生成模型 Nano Banana 2。它提供先進的世界知識、生產就緒規格、主體一致性等功能,全以閃電般的 Flash 速度運行。
2月20日,微軟CEO薩提亞·納德拉宣布將xAI的Grok 4.1 Fast加入多模型產品系列。這擴展了開發者在Azure AI服務中的模型選擇。

用戶讚揚 Z.AI 私有模型 GLM-5-Turbo,在 OpenRouter 測試速度極快、智慧達 Gemini 3.2 Flash 水平或更佳。尚未上 Hugging Face。盼望未來開源。

OpenAI released GPT-5.3-Codex-Spark, a lightweight version of its Codex intelligent agent programming tool. This slimmed-down model prioritizes extreme inference speed for rapid iteration scenarios. It follows the latest full Codex model released earlier this month.

OpenAI released GPT-5.3-Codex-Spark, a lightweight version of its Codex programming tool. It targets rapid iteration with extreme inference speed. This follows a recent full Codex model update.