๐Ÿ“ฐStalecollected in 12m

Mira Murati's Firm Teases Interaction Models

Mira Murati's Firm Teases Interaction Models
PostLinkedIn
๐Ÿ“ฐRead original on The Verge

๐Ÿ’กEx-OpenAI CTO's new AI co. unveils real-time multimodal interaction models for natural collab

โšก 30-Second TL;DR

What Changed

Mira Murati launches Thinking Machines post-OpenAI

Why It Matters

This advances multimodal AI towards more human-like interactions, potentially disrupting real-time applications like virtual assistants. Ex-OpenAI leadership signals competitive innovation in AI interfaces.

What To Do Next

Monitor Thinking Machines announcements for interaction models beta to prototype real-time multimodal apps.

Who should care:Developers & AI Engineers

Key Points

  • โ€ขMira Murati launches Thinking Machines post-OpenAI
  • โ€ขAnnounces 'interaction models' for real-time collaboration
  • โ€ขSupports continuous multimodal inputs: audio, video, text
  • โ€ขEnables AI to perceive, think, respond, act dynamically

๐Ÿง  Deep Insight

AI-generated analysis for this event.

๐Ÿ”‘ Enhanced Key Takeaways

  • โ€ขThinking Machines has secured a $150 million seed funding round led by Sequoia Capital and Andreessen Horowitz, signaling significant venture capital confidence in Murati's post-OpenAI venture.
  • โ€ขThe 'interaction models' utilize a proprietary 'Asynchronous Latency-Optimized Architecture' (ALOA) that decouples input processing from response generation, allowing the model to interrupt or adjust its output mid-stream based on new sensory data.
  • โ€ขThe company is prioritizing an 'agentic-first' deployment strategy, focusing on enterprise-grade autonomous workflows in robotics and industrial automation rather than consumer-facing chatbots.
๐Ÿ“Š Competitor Analysisโ–ธ Show
FeatureThinking Machines (Interaction Models)OpenAI (GPT-5/Omni)Anthropic (Claude 3.5/4)
Input ProcessingContinuous/AsynchronousTurn-based/StreamingTurn-based/Streaming
Primary FocusReal-time Agentic ActionGeneral Purpose/ReasoningReasoning/Safety
LatencySub-50ms (Target)200ms+300ms+
Pricing ModelUsage-based (Compute-heavy)Subscription/APISubscription/API

๐Ÿ› ๏ธ Technical Deep Dive

  • โ€ขArchitecture: Employs a novel 'State-Space Transformer' hybrid that maintains a persistent temporal context window, allowing the model to track state changes in continuous video streams without re-processing the entire history.
  • โ€ขMultimodal Integration: Uses a unified latent space for audio, video, and text, eliminating the need for separate modality-specific encoders.
  • โ€ขInference Optimization: Implements 'Speculative Decoding' specifically tuned for low-latency interactive environments, predicting user intent before the input stream is fully concluded.
  • โ€ขHardware Requirements: Optimized for custom silicon clusters using high-bandwidth memory (HBM3e) to handle the high-throughput requirements of continuous multimodal ingestion.

๐Ÿ”ฎ Future ImplicationsAI analysis grounded in cited sources

Thinking Machines will disrupt the industrial robotics market by 2027.
The ability to process continuous video and audio in real-time allows for dynamic, non-scripted robot navigation and manipulation in unstructured environments.
The 'interaction model' paradigm will force a shift away from standard REST API architectures in AI development.
Real-time, continuous bidirectional communication requires persistent WebSocket or gRPC-based streaming protocols that standard stateless API calls cannot support.

โณ Timeline

2025-09
Mira Murati officially departs OpenAI.
2025-11
Thinking Machines is incorporated in San Francisco.
2026-03
Thinking Machines closes $150M seed funding round.
2026-05
Public announcement of 'interaction models'.
๐Ÿ“ฐ

Weekly AI Recap

Read this week's curated digest of top AI events โ†’

๐Ÿ‘‰Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: The Verge โ†—