StreamingMemeStreamingMemeBuyers Guide
AboutLeaderboardsEventsSubmit News
Subscribe

Daily Brief

The streaming industry in your inbox every morning.

Daily Brief

The streaming industry in your inbox every morning.

StreamingMemeStreamingMeme

The independent buyers guide and news aggregator for the streaming technology industry.

Explore

Buyers GuideLeaderboardsEventsSubmit News

Stay updated

Weekly digest of new companies and streaming news.

Categories

Encoding & SoftwareVideo Delivery & CDNStreaming PlatformsAI for VideoProduction HardwareBusiness NewsMonetization & Ad TechRegulatory & Policy

© 2026 StreamingMeme. All rights reserved.

AboutPrivacy PolicyTermsContact
EncodingCDNPlatformsAI & VideoHardwareBusinessAd TechPolicy
← AI for Video
AI & VideoProduct LaunchJuly 9, 2026

D-Matrix launches Corsair accelerators to monetizing premium fast tokens

D-Matrix launches Corsair accelerators to monetizing premium fast tokens
SiliconANGLE Media

d-Matrix is launching its Corsair accelerators in a heterogeneous partnership with Nvidia to improve token generation speeds for AI inference. The deployment aims to solve memory bandwidth bottlenecks in real-time agentic AI, enabling the monetization of low-latency, high-interactivity 'fast tokens' for streaming and application developers.

Key Takeaways

  • Commercial deployment with Parasail pairs d-Matrix Corsair cards with Nvidia Hopper and Blackwell GPUs.
  • Corsair platform achieves 10x faster inference and 3x lower costs by disaggregating prefill and decode tasks.
  • Integrated 3D memory architecture stacks DRAM and logic to bypass standard high-bandwidth memory (HBM) limits.
  • Fast tokens enable new revenue tiers, with developers charging up to 10x premiums for real-time interactivity.

Why It Matters

The shift toward agentic AI is exposing the 'memory wall' where GPU compute speeds outpace data transfer rates. This heterogeneous approach moves beyond the GPU-only era, positioning specialized accelerators as essential companions rather than competitors to Nvidia’s dominance. For the streaming industry, this technology provides the infrastructure necessary to scale low-latency AI agents and high-interactivity video features that require immediate response speeds. As inference moves from research to production, the ability to monetize 'fast tokens' will determine the profitability of next-generation AI services. Watch for hyperscaler adoption rates of disaggregated inference racks to signal a broader shift in data center architecture.

Additional Context

The production launch of the Corsair platform in June 2026 follows a major $275 million Series C round in late 2025 led by Temasek and Microsoft's M12, which valued d-Matrix at $2 billion. Per AIWeekly (July 2026), the company is shipping to a mix of hyperscalers and 'neoclouds' that are looking to offset the high total cost of ownership associated with standalone Nvidia Blackwell clusters. This trend toward specialization is reflected across the sector; per New Market Pitch (June 2026), startups like Etched and MatX are similarly raising hundreds of millions to develop ASIC-based alternatives to general-purpose GPUs. Market demand is largely driven by the 'Fast Mode' capabilities in frontier models. Anthropic, for example, updated its pricing on the Claude API in May 2026 to include a high-speed tier for its Opus 4.8 model. Per Anthropic's official documentation (June 2026), this fast mode delivers responses up to 2.5x faster at approximately double the standard token price. Such pricing models confirm d-Matrix's thesis that latency has become a premium commodity that enterprises are willing to pay for to enable real-time collaborative coding and agentic workflows. Technical bottlenecks remain a critical industry focus. According to Qualcomm (July 2026), memory-bound workloads are increasingly exposing the limits of conventional XPU and HBM architectures, leading to a rise in 'near-memory' computing solutions. By placing computation directly on the memory substrate, as d-Matrix does with its 6,400 mm² silicon cards, providers can achieve bandwidth reaching 300 TB/s. This is significant because, per Spheron (April 2026), the decode phase of large language models sits well below the performance roofline of traditional GPUs due to low arithmetic intensity.


Read full article at siliconangle.com

Get this in your inbox → Subscribe

Enjoy our coverage?

Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.

Add as preferred source

Related Articles

Content+Technology: Runway launches Media Router to automate generative video model selection
Wccftech: Qualcomm Adreno 850 GPU to debut AI Frame Fusion technology
X: vLLM v0.26.0 introduces tiered KV offloading and multimodal audio-video support

Newest

about 5 hours ago
Cord Cutters News: Paramount recruits veteran Microsoft defense attorney to fight California merger block
1 day ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
1 day ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
1 day ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
1 day ago
Investing.com: TF1 Digital Revenues Jump 17% as Netflix Partnership Exceeds Growth Targets
1 day ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
1 day ago
YouTube: Microsoft tests ad-supported Xbox Cloud Gaming tier for Xbox Insiders
1 day ago
SatNews: FCC proposes unlicensed 2.4 GHz spectrum for direct-to-satellite IoT links
1 day ago
Beet.TV: Brands must re-describe catalogs for AI agents to maintain discoverability
1 day ago
Wilkinson Barker Knauer LLP: FCC orders Upper C-band spectrum clearing as ATSC 3.0 reaches top markets
1 day ago
Yahoo: LG mandates removal of residential proxy SDKs from webOS apps
1 day ago
Wccftech: Qualcomm Adreno 850 GPU to debut AI Frame Fusion technology
1 day ago
Content+Technology: Runway launches Media Router to automate generative video model selection
1 day ago
Associated Press: Moonshot Kimi K3 leads surge of Chinese AI adoption in U.S.
1 day ago
daily.dev: AVIF achieves universal browser support as Edge and Safari close gaps
1 day ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
1 day ago
Los Angeles Times: Disney, Netflix, and Amazon recruit AI talent to automate production workflows
1 day ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
1 day ago
Callaba: Callaba standardizes remote production workflows via SRT and NDI integration
1 day ago
AI Rights Brief: Google and Disney integrate AI provenance directly into programmatic ad workflows

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.YouTube63
  4. 4.Tech Times60
  5. 5.AdExchanger57
  6. 6.TechCrunch55
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →

Newest

about 5 hours ago
Cord Cutters News: Paramount recruits veteran Microsoft defense attorney to fight California merger block
1 day ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
1 day ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
1 day ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
1 day ago
Investing.com: TF1 Digital Revenues Jump 17% as Netflix Partnership Exceeds Growth Targets
1 day ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
1 day ago
YouTube: Microsoft tests ad-supported Xbox Cloud Gaming tier for Xbox Insiders
1 day ago
SatNews: FCC proposes unlicensed 2.4 GHz spectrum for direct-to-satellite IoT links
1 day ago
Beet.TV: Brands must re-describe catalogs for AI agents to maintain discoverability
1 day ago
Wilkinson Barker Knauer LLP: FCC orders Upper C-band spectrum clearing as ATSC 3.0 reaches top markets
1 day ago
Yahoo: LG mandates removal of residential proxy SDKs from webOS apps
1 day ago
Wccftech: Qualcomm Adreno 850 GPU to debut AI Frame Fusion technology
1 day ago
Content+Technology: Runway launches Media Router to automate generative video model selection
1 day ago
Associated Press: Moonshot Kimi K3 leads surge of Chinese AI adoption in U.S.
1 day ago
daily.dev: AVIF achieves universal browser support as Edge and Safari close gaps
1 day ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
1 day ago
Los Angeles Times: Disney, Netflix, and Amazon recruit AI talent to automate production workflows
1 day ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
1 day ago
Callaba: Callaba standardizes remote production workflows via SRT and NDI integration
1 day ago
AI Rights Brief: Google and Disney integrate AI provenance directly into programmatic ad workflows

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.YouTube63
  4. 4.Tech Times60
  5. 5.AdExchanger57
  6. 6.TechCrunch55
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →