▲Vercel News•較早收集於 7h
GPT 5.3 Codex 現已在 AI Gateway 上線

#ai-gateway#agentic-ai#model-efficiencyvercel-ai-gatewayvercelopenaigpt-5.3-codex
💡25% faster Codex model for agentic coding now on Vercel Gateway—boost your dev workflows
⚡ 30-Second TL;DR
有什麼變化
融合 GPT-5.2-Codex 編碼與 GPT-5.2 推理能力
為什麼重要
此更新透過更快、具脈絡意識的編碼代理,提升開發者效率,並在可靠閘道上降低成本,藉由 token 效率與重試等優化。將 Vercel 定位為生產 AI 編碼工作流程的關鍵基礎設施。
下一步行動
Set model to 'openai/gpt-5.3-codex' in your Vercel AI SDK to test agentic coding tasks immediately.
誰應關注:Developers & AI Engineers
關鍵要點
- •融合 GPT-5.2-Codex 編碼與 GPT-5.2 推理能力
- •比前代快 25% 且更節省 token
- •支援長時間代理式任務如研究、工具使用、多步驟執行
- •改善網頁開發對不完整提示的處理,提供生產就緒輸出
- •透過 Vercel AI SDK 中的 'openai/gpt-5.3-codex' 存取
🧠 深度解析
背景與延伸:來自公開資料,非原文內容。引用 7 個來源。
🔑 增強重點摘要
- •GPT-5.3-Codex-Spark, a smaller optimized variant, was released in research preview on February 12, 2026, delivering over 1,000 tokens/second via Cerebras Wafer-Scale Engine hardware for real-time coding with near-instant feedback[6][7]
- •The model achieves state-of-the-art performance on SWE-Bench Pro (spanning four programming languages) and Terminal-Bench 2.0, while nearly doubling its OSWorld-Verified benchmark score compared to predecessors[4][5]
- •GPT-5.3-Codex expanded beyond pure coding to handle end-to-end professional workflows including Jira ticket updates, documentation generation, deployment pipeline management, and cybersecurity tasks with 'High capability' rating[5]
- •The model was optimized for NVIDIA GB200 NVL72 hardware and employs conversation compaction techniques to efficiently manage 1M token context windows in agentic loops[5]
- •GitHub Copilot integrated GPT-5.3-Codex on February 9, 2026, making it available across Copilot Pro, Pro+, Business, and Enterprise tiers in Visual Studio Code, GitHub Mobile, CLI, and Coding Agent[3]
🛠️ 技術深入
- •Architecture: Merges frontier coding performance of GPT-5.2-Codex with reasoning and professional knowledge capabilities of GPT-5.2 into a unified model[1][4]
- •Inference Optimization: 25% faster than GPT-5.2-Codex through infrastructure improvements and optimized inference stack; achieves higher accuracy with fewer tokens[1][4][5]
- •Hardware Optimization: Optimized for NVIDIA GB200 NVL72 to reduce latency in agentic loops; Codex-Spark variant runs on Cerebras Wafer-Scale Engine at 1,000+ tokens/second[5][6][7]
- •Context Management: 1M token context window with conversation compaction for efficient long-history management in multi-step workflows[5]
- •Benchmark Performance: SWE-Bench Pro (state-of-the-art across 4 languages), Terminal-Bench 2.0 (75.1% accuracy), OSWorld-Verified (nearly doubled score), GDPval[4][5]
- •Regression Fixes: Reduced non-deterministic linting loops, improved bug-analysis evidence quality, lowered premature completion in flaky-test scenarios[1]
🔮 前景展望AI analysis grounded in cited sources
Agentic autonomy will expand beyond software engineering into enterprise operations
Self-healing infrastructure and legacy migration capabilities suggest GPT-5.3-Codex enables autonomous agents to manage production systems, code rewrites, and documentation simultaneously without human intervention[5]
Real-time coding workflows will fragment into two model classes: frontier (long-running tasks) and Spark (instant feedback)
The dual-model strategy with Codex-Spark optimized for sub-second latency indicates the market is bifurcating between ambitious multi-day agentic tasks and interactive development requiring immediate responsiveness[7]
Cybersecurity automation will accelerate through high-capability vulnerability detection
GPT-5.3-Codex's 'High capability' cybersecurity rating and direct vulnerability detection enable automated penetration testing and patching at scale, reducing manual security review bottlenecks[5]
⏳ 時間線
2026-01
OpenAI announces partnership with Cerebras for ultra-low latency inference
2026-02-05
GPT-5.3-Codex launches across all Codex surfaces (app, CLI, IDE extension, web) for paid ChatGPT users; API access announced for coming weeks
2026-02-09
GPT-5.3-Codex becomes generally available in GitHub Copilot for Pro, Pro+, Business, and Enterprise users with gradual rollout
2026-02-12
GPT-5.3-Codex-Spark research preview released, powered by Cerebras, delivering 1,000+ tokens/second for real-time coding
📎 來源 (7)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
- digitalapplied.com — Gpt 5 3 Codex Release Features Benchmarks Guide
- community.openai.com — 1373453
- github.blog — 2026 02 09 Gpt 5 3 Codex Is Now Generally Available for Github Copilot
- OpenAI — Introducing Gpt 5 3 Codex
- datacamp.com — Gpt 5 3 Codex
- cerebras.ai — Openai Codexspark
- OpenAI — Introducing Gpt 5 3 Codex Spark
📰
AI 週報
閱讀本週精選 AI 大事摘要 →
👉相關動態
AI 策展新聞聚合。所有內容版權歸原始發布者所有。
原始來源: Vercel News ↗
每週 AI 簡報
每週一封,可隨時退訂。