StreamingMemeStreamingMemeBuyers Guide
AboutLeaderboardsEventsSubmit News
Subscribe

Daily Brief

The streaming industry in your inbox every morning.

Daily Brief

The streaming industry in your inbox every morning.

StreamingMemeStreamingMeme

The independent buyers guide and news aggregator for the streaming technology industry.

Explore

Buyers GuideLeaderboardsEventsSubmit News

Stay updated

Weekly digest of new companies and streaming news.

Categories

Encoding & SoftwareVideo Delivery & CDNStreaming PlatformsAI for VideoProduction HardwareBusiness NewsMonetization & Ad TechRegulatory & Policy

© 2026 StreamingMeme. All rights reserved.

AboutPrivacy PolicyTermsContact
EncodingCDNPlatformsAI & VideoHardwareBusinessAd TechPolicy
← AI for Video
AI & VideoTechnical DevelopmentJune 4, 2026

Mira Murati previews Thinking Machines' continuous-stream multimodal AI architecture

Mira Murati previews Thinking Machines' continuous-stream multimodal AI architecture
Let's Data Science

Mira Murati, co-founder and CEO of Thinking Machines Lab, presented their new 'interaction models' approach at Bloomberg Tech. These models process audio, video, and text as continuous, parallel streams for real-time, multimodal interaction, diverging from traditional turn-based systems. The approach requires low-latency streaming inference and integrated capture, posing significant compute and engineering challenges.

Key Takeaways

  • Thinking Machines' architecture processes 200ms of input while generating 200ms of output to achieve sub-400ms total response latency.
  • The lab's only shipping product as of June 2026 is Tinker, an API for fine-tuning open-source models that launched in October 2025.
  • Murati confirmed the startup is utilizing 128K context windows and mixture-of-experts (MoE) backbones for its interaction models.
  • The startup has raised $2 billion to date from investors including Andreessen Horowitz and Nvidia to fund compute-heavy streaming inference.

Why It Matters

Native multimodal streaming represents a shift from the 'bolted-on' interactivity of current voice assistants to an architecture where latency is a core model feature. For the streaming industry, this suggests a future where AI can monitor live video feeds—such as sports or security—and provide mid-stream feedback without the 1-2 second lag typical of batch processing. The challenge remains the extreme compute cost of maintaining persistent state for continuous video and audio inputs at under 200ms granularity. Success will depend on whether Thinking Machines can move beyond the 'Tinker' development wedge into production-ready deployments as rivals like OpenAI and Google iterate on their own low-latency voice and vision modes.

Additional Context

Since Murati’s departure from OpenAI in September 2024, the competitive landscape for low-latency multimodal interaction has narrowed to a high-stakes hardware race. Per Bloomberg (June 2026), Thinking Machines Lab reached a $12 billion valuation following a $2 billion seed round led by Andreessen Horowitz. To support the heavy compute demands of parallel audio-video streams, the startup secured a multiyear supply agreement with Nvidia for its Vera Rubin GPU architecture in March 2026, according to Crypto Briefing. This hardware commitment is critical as the lab aims to scale its TML-Interaction-Small model for public release by late 2026. While Thinking Machines emphasizes transparency, it faces significant internal volatility. Per The Next Web (June 2026), the lab has navigated a series of researcher departures to Meta and OpenAI despite its aggressive hiring from those same firms during its 2025 launch phase. Observers such as Puck News (May 2026) noted that until the Bloomberg appearance, the company had been 'eerily quiet,' relying on Tinker—a service for fine-tuning open-weights models like Llama 3—to maintain developer mindshare while focusing internally on the higher-complexity streaming inference stack. Technically, the 'interaction model' approach mirrors a broader industry transition toward full-duplex systems. Competitors such as Kyutai’s Moshi model previously demonstrated theoretical latencies as low as 160ms for audio, according to Unitlab (January 2026). However, Murati’s vision adds continuous video processing to the mix, requiring significantly higher throughput. As OpenAI prepares for a reported IPO and Anthropic targets a $1 trillion valuation, Thinking Machines is positioning its 'human-AI collaboration' thesis as a strategic alternative to the increasingly closed ecosystems of incumbent frontier labs.


Read full article at letsdatascience.com

Get this in your inbox → Subscribe

Enjoy our coverage?

Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.

Add as preferred source

Related Articles

MarkTechPost: Induction Labs Photon-1 trains on 18 years of raw video
MarkTechPost: Reactor releases 1.6B parameter open-source Dreamer 4 world-model implementation
YouTube: NTT's LLMlet enables distributed LLM inference across browsers via WebRTC

Newest

2 days ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
2 days ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
2 days ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
2 days ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
2 days ago
Front Office Sports: World Cup afternoon ratings spark shift toward earlier U.S. game windows
2 days ago
Cord Cutters News: FCC chair signals scrutiny for potential streaming-exclusive 2030 World Cup rights
2 days ago
Cord Cutters News: Linear contraction accelerates as 14 cable networks vanish in five years
2 days ago
Los Angeles Times: Disney, Netflix, and Amazon recruit AI talent to automate production workflows
2 days ago
Callaba: Callaba standardizes remote production workflows via SRT and NDI integration
2 days ago
Wccftech: Qualcomm Adreno 850 GPU to debut AI Frame Fusion technology
2 days ago
AI Rights Brief: Google and Disney integrate AI provenance directly into programmatic ad workflows
2 days ago
Lib.rs: New zero-dependency Rust decoder vp9dec achieves bit-exact VP9 conformance
2 days ago
Futurism: Meta and TikTok face backlash over deceptive AI-generated health ads
2 days ago
Euronews: EU Expert Panel Backs Age Restrictions and Addictive Feature Bans
2 days ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
2 days ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
2 days ago
IPWatchdog: EC mandates Google share search data and Android features under DMA
2 days ago
TechRadar: Weka's new WEKApod 3 uses Micron 245TB SSDs for exabyte-scale storage
2 days ago
Lib.rs: Moq-relay 0.3.1 adds mTLS and admission policies for production-grade QUIC streaming
2 days ago
Yahoo: LG mandates removal of residential proxy SDKs from webOS apps

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.YouTube63
  4. 4.Tech Times60
  5. 5.AdExchanger57
  6. 6.TechCrunch55
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →

Newest

2 days ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
2 days ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
2 days ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
2 days ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
2 days ago
Front Office Sports: World Cup afternoon ratings spark shift toward earlier U.S. game windows
2 days ago
Cord Cutters News: FCC chair signals scrutiny for potential streaming-exclusive 2030 World Cup rights
2 days ago
Cord Cutters News: Linear contraction accelerates as 14 cable networks vanish in five years
2 days ago
Los Angeles Times: Disney, Netflix, and Amazon recruit AI talent to automate production workflows
2 days ago
Callaba: Callaba standardizes remote production workflows via SRT and NDI integration
2 days ago
Wccftech: Qualcomm Adreno 850 GPU to debut AI Frame Fusion technology
2 days ago
AI Rights Brief: Google and Disney integrate AI provenance directly into programmatic ad workflows
2 days ago
Lib.rs: New zero-dependency Rust decoder vp9dec achieves bit-exact VP9 conformance
2 days ago
Futurism: Meta and TikTok face backlash over deceptive AI-generated health ads
2 days ago
Euronews: EU Expert Panel Backs Age Restrictions and Addictive Feature Bans
2 days ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
2 days ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
2 days ago
IPWatchdog: EC mandates Google share search data and Android features under DMA
2 days ago
TechRadar: Weka's new WEKApod 3 uses Micron 245TB SSDs for exabyte-scale storage
2 days ago
Lib.rs: Moq-relay 0.3.1 adds mTLS and admission policies for production-grade QUIC streaming
2 days ago
Yahoo: LG mandates removal of residential proxy SDKs from webOS apps

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.YouTube63
  4. 4.Tech Times60
  5. 5.AdExchanger57
  6. 6.TechCrunch55
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →