Meta、Muse Spark 1.3でコーディング性能を強化
💡A new coding-and-agent model claims a major leap and frontier-level DeepSWE performance.
⚡ 30-Second TL;DR
What Changed
Improves coding performance and AI agent capabilities.
Why It Matters
The release could raise expectations for models that independently execute complex coding workflows over extended periods. API availability and a planned open-weight release may give developers both a production integration path and a more flexible experimentation option.
What To Do Next
Evaluate Muse Spark 1.3 through its API on a long-running coding workflow and compare completion quality against your current model using DeepSWE-style tasks.
Key Points
- •Improves coding performance and AI agent capabilities.
- •Handles ambiguous instructions and maintains progress during longer-running tasks.
- •Outperformed other frontier models on benchmarks including DeepSWE.
- •Available through APIs and other channels, with the previous model planned for open-weight release.
🧠 Deep Insight
Background and context from public sources — not the original article. 11 sources cited.
🔑 Enhanced Key Takeaways
- •Muse Spark 1.3 achieved a score of 75.4 on the DeepSWE v1.1 benchmark, specifically outperforming GPT-5.6 Sol and Claude Opus 5.
- •The model demonstrates superior long-context comprehension, scoring 98.1 on the MRCR benchmark within the 512K to 1M token range.
- •Operational efficiency has been significantly improved over the 1.2 version, with a 20% reduction in tool-calling overhead and a 25% decrease in token consumption.
- •The model includes enhanced safety protocols, specifically increasing resistance to adversarial inputs and prompt injection attacks during autonomous operations.
- •Deployment is facilitated through the 'Muse Code' developer platform and the Meta Model API, with immediate availability following the September 2, 2026 announcement.
📊 Competitor Analysis▸ Show
| Feature | Muse Spark 1.3 | GPT-5.6 Sol | Claude Opus 5 |
|---|---|---|---|
| DeepSWE v1.1 Score | 75.4 | < 75.4 | < 75.4 |
| Context Window | 1M+ tokens | N/A | N/A |
| Efficiency | 25% token reduction | Standard | Standard |
| Pricing | API-based | Premium | Premium |
🛠️ Technical Deep Dive
- Optimized for long-running coding tasks with improved autonomous error recovery and intent clarification logic.
- Architecture features refined tool-calling mechanisms that reduce redundant API interactions by 20%.
- Enhanced safety layer specifically tuned to prevent irreversible system operations without user verification.
- High-performance long-context processing capability validated up to 1M tokens via the MRCR benchmark.
🔮 Future ImplicationsAI analysis grounded in cited sources
⏳ Timeline
📎 Sources (11)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
Weekly AI Recap
Read this week's curated digest of top AI events →
👉Related Updates
AI-curated news aggregator. All content rights belong to original publishers.
Original source: ITmedia AI+ (日本) ↗
This is a summary, not the original. Read the source, or get the weekly briefing.
Weekly AI briefing
One email a week. Unsubscribe anytime.


