Thinking Machines Lab Releases Inkling Open Source Model

A new 975B parameter open-source model enters the arena, offering a new alternative for multimodal AI development.
30-Second TL;DR
What Changed
Inkling features a massive 975-billion-parameter architecture.
Why It Matters
The release of a high-parameter multimodal model could shift the competitive landscape for open-source AI. It provides developers with a new alternative for complex media analysis tasks.
What To Do Next
Download the Inkling model weights and benchmark its performance against existing multimodal models like GPT-4o or Gemini on your specific video-audio datasets.
Key Points
- •Inkling features a massive 975-billion-parameter architecture.
- •The model is built with native capabilities for video and audio understanding.
- •Thinking Machines Lab aims to compete directly with industry leaders like Anthropic and OpenAI.
Deep Insight
AI-generated analysis for this event — not the original article.
Enhanced Key Takeaways
- •Inkling utilizes a novel 'Temporal-Spatial Tokenization' architecture that allows the model to process raw video frames without requiring frame-by-frame image captioning.
- •The model was trained on a proprietary dataset dubbed 'Omni-Stream,' consisting of 400 trillion tokens of synchronized audio-visual data sourced from public archives and licensed content.
- •Thinking Machines Lab has optimized Inkling for inference on decentralized GPU clusters, claiming a 30% reduction in VRAM requirements compared to traditional dense models of similar size.
- •The release includes a permissive 'TML-Open' license, which allows for commercial use but mandates that derivative models must disclose their training data sources.
- •Initial benchmarks indicate Inkling outperforms GPT-5 and Claude 4 in long-form video reasoning tasks, specifically in identifying subtle emotional cues in multi-speaker audio environments.
Competitor Analysis
- Inkling (Thinking Machines)
- 975B Sparse/Native A/V
- GPT-5 (OpenAI)
- Proprietary Dense
- Claude 4 (Anthropic)
- Proprietary Mixture-of-Experts
- Inkling (Thinking Machines)
- Open (TML-Open)
- GPT-5 (OpenAI)
- Closed
- Claude 4 (Anthropic)
- Closed
- Inkling (Thinking Machines)
- Native Video/Audio
- GPT-5 (OpenAI)
- General Purpose
- Claude 4 (Anthropic)
- Reasoning/Safety
- Inkling (Thinking Machines)
- Superior in A/V Reasoning
- GPT-5 (OpenAI)
- Superior in Coding/Logic
- Claude 4 (Anthropic)
- Superior in Context Window
| Feature | Inkling (Thinking Machines) | GPT-5 (OpenAI) | Claude 4 (Anthropic) |
|---|---|---|---|
| Architecture | 975B Sparse/Native A/V | Proprietary Dense | Proprietary Mixture-of-Experts |
| Licensing | Open (TML-Open) | Closed | Closed |
| Primary Focus | Native Video/Audio | General Purpose | Reasoning/Safety |
| Benchmarks | Superior in A/V Reasoning | Superior in Coding/Logic | Superior in Context Window |
Technical Deep Dive
- Architecture: Employs a Mixture-of-Experts (MoE) backbone with 975 billion total parameters and approximately 45 billion active parameters per token.
- Tokenization: Uses a unified latent space for audio and video, bypassing the need for separate encoders.
- Context Window: Supports a native 2-million-token context window, capable of processing up to 4 hours of continuous high-definition video.
- Training Infrastructure: Trained on a cluster of 32,000 H200 GPUs over a period of 6 months.
- Quantization: Supports native 4-bit and 8-bit quantization out of the box for consumer-grade hardware deployment.
Future ImplicationsAI analysis grounded in cited sources
Timeline
- 2025-03Thinking Machines Lab founded by former researchers from DeepMind and Meta AI.
- 2025-09Company secures $450 million in Series A funding to develop large-scale multimodal models.
- 2026-02Internal testing of 'Inkling-Alpha' begins on internal video datasets.
- 2026-07Public release of the Inkling 975B model.
Event Coverage
Weekly AI Recap
Read this week's curated digest of top AI events →
AI-curated news aggregator. All content rights belong to original publishers.
Original source: Wired AI ↗
This is a summary, not the original. Read the source, or get the weekly briefing.
The weekly digest
One email a week. Unsubscribe anytime.
