Openrouter Hunter/Healer Confirmed as MiMo V2
๐ก1M token context + image model confirmed on Openrouter โ ideal for long RAG tasks!
โก 30-Second TL;DR
What Changed
Hunter Alpha = MiMo V2 Pro: 1M (1,048,576) token context, 32K max output, text-only
Why It Matters
Provides AI practitioners with accessible high-context, multimodal models on Openrouter, enabling advanced RAG and long-document analysis without custom hosting.
What To Do Next
Test MiMo V2 Pro on Openrouter API for 1M context reasoning benchmarks.
Key Points
- โขHunter Alpha = MiMo V2 Pro: 1M (1,048,576) token context, 32K max output, text-only
- โขHealer Alpha = MiMo V2 Omni: 262K context, text+image reasoning, 32K max output
- โขConfirmation via openclaw/openclaw Github PR #49214
- โขNew unidentified model announced as incoming
๐ง Deep Insight
Background and context from public sources โ not the original article. 8 sources cited.
๐ Enhanced Key Takeaways
- โขMiMo-V2-Flash, a related Xiaomi model, is a 309B total parameter Mixture-of-Experts (MoE) architecture with 15B active parameters and hybrid attention, released on December 14, 2025[1][2].
- โขMiMo-V2-Flash leads global open-source models on SWE-bench Verified and SWE-bench Multilingual benchmarks, matching Claude Sonnet 4.5 performance at 3.5% of the cost[1][2].
- โขMiMo-V2-Flash pricing is $0.09 per 1M input tokens with 262K context and 16K output limit via OpenRouter integration[3].
- โขThe model supports a 'reasoning: enabled' toggle for step-by-step thinking, with reasoning_details in responses for agent workflows[1][6].
๐ Competitor Analysisโธ Show
| Model | Provider | Context Window | Pricing (Input/1M tokens) | Key Benchmarks |
|---|---|---|---|---|
| MiMo-V2-Flash | Xiaomi | 262K | $0.09 | #1 open-source SWE-bench Verified/Multilingual[1][2][3] |
| Devstral 2 | Mistral | 262K | Free | Strong SWE-Bench coding[5] |
| Gemini 2.0 Flash Exp | 1M | Free | Long documents, multimodal[5] | |
| Qwen3-Coder | Qwen | 262K | Free | Strong code reasoning[5] |
๐ ๏ธ Technical Deep Dive
- โขMiMo-V2-Flash employs Mixture-of-Experts (MoE) with 309B total parameters and 15B active parameters per inference[1][2].
- โขFeatures hybrid attention architecture and a hybrid-thinking toggle controllable via 'reasoning: enabled' boolean parameter[1][2][6].
- โขSupports 262,144 token context window, text input, and parameters like frequency_penalty, temperature, tools, and tool_choice[4].
- โขOpenRouter integration exposes reasoning_details array in responses for preserving step-by-step reasoning across conversations[6].
๐ฎ Future ImplicationsAI analysis grounded in cited sources
โณ Timeline
๐ Sources (8)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
Weekly AI Recap
Read this week's curated digest of top AI events โ
๐Related Updates
AI-curated news aggregator. All content rights belong to original publishers.
Original source: Reddit r/LocalLLaMA โ
This is a summary, not the original. Read the source, or get the weekly briefing.
Weekly AI briefing
One email a week. Unsubscribe anytime.


