StreamingMemeStreamingMemeBuyers Guide
AboutLeaderboardsEventsSubmit News
Subscribe

Daily Brief

The streaming industry in your inbox every morning.

Daily Brief

The streaming industry in your inbox every morning.

StreamingMemeStreamingMeme

The independent buyers guide and news aggregator for the streaming technology industry.

Explore

Buyers GuideLeaderboardsEventsSubmit News

Stay updated

Weekly digest of new companies and streaming news.

Categories

Encoding & SoftwareVideo Delivery & CDNStreaming PlatformsAI for VideoProduction HardwareBusiness NewsMonetization & Ad TechRegulatory & Policy

© 2026 StreamingMeme. All rights reserved.

AboutPrivacy PolicyTermsContact
EncodingCDNPlatformsAI & VideoHardwareBusinessAd TechPolicy
← AI for Video
AI & VideoFunding RoundJuly 20, 2026

Infinity raises $15M to breach Nvidia’s CUDA moat with automated code

Infinity raises $15M to breach Nvidia’s CUDA moat with automated code
TechCrunch

Inference startup Infinity has raised $15 million at a $100 million valuation to develop its Ignition agent, which automatically writes and optimizes kernel code for AI inference across diverse chip architectures. The software aims to reduce reliance on the Nvidia CUDA ecosystem by enabling high-performance AI deployment on alternative hardware.

Key Takeaways

  • Seed round led by Touring Capital with participation from OpenAI and Anthropic researchers.
  • Ignition agent uses a self-optimizing feedback loop to write, test, and debug low-level chip code.
  • Early case study showed the agent reached 92% of d-Matrix’s Corsair chip peak performance in 10 hours.
  • Business model avoids upfront licensing, instead taking a cut of cost savings and performance gains.
  • Revenue is already being generated through a ship-design partnership with hardware challenger d-Matrix.

Why It Matters

The standard barrier for non-Nvidia hardware has always been the software gap, as writing high-performance kernels for new chips can take years of human engineering. Infinity’s automated approach collapses this timeline to days, potentially commoditizing AI hardware by making migration between chips friction-free. For the streaming industry, which is pivotally shifting from model training to high-volume inference, this provides a path toward significantly lower token costs and reduced dependence on tight GPU supply chains. Watch for rival silicon providers to adopt similar automated stacks to accelerate time-to-market for their specialized inference accelerators.

Additional Context

The funding for Infinity arrives as the AI sector undergoes a structural shift from model training to large-scale inference. According to industry estimates cited by Business Wire in July 2026, inference workloads are projected to represent roughly two-thirds of all AI compute spending this year. This shift has intensified the search for alternatives to Nvidia’s dominant H100 and H200 GPUs. Per research from TrendForce in June 2026, shipments for custom AI Application-Specific Integrated Circuits (ASICs) are expected to grow 44.6% in 2026, significantly outpacing the 16.1% growth projected for general-purpose GPUs. Hardware challengers are already moving toward full-scale deployment to meet this demand. For instance, d-Matrix, a key partner for Infinity, announced in June 2026 that its Corsair inference platform has entered volume production. D-Matrix claims its Corsair chips deliver 10x faster performance and 5x better energy efficiency for Large Language Model (LLM) inference than traditional GPUs, specifically targeting the latency-sensitive needs of real-time voice and agentic AI applications. Per AI Multiple, d-Matrix reached a $2 billion valuation following a $275 million Series C round led by Temasek and Microsoft's M12 in late 2025. While hardware performance improves, the software remains the primary bottleneck for widespread adoption. Per Million Miner in July 2026, Nvidia's CUDA ecosystem still maintains approximately six million developers and 18 years of deeply integrated libraries. Competitors like AMD have tried to close this gap with the ROCm platform, and its MI355X chip reportedly matches Nvidia's Blackwell on raw compute. However, the introduction of automated agents like Infinity’s Ignition marks a new strategy: using AI itself to bypass the manual labor of software porting, which has historically been the most effective protector of Nvidia's 80% market share.


Read full article at techcrunch.com

Get this in your inbox → Subscribe

Enjoy our coverage?

Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.

Add as preferred source

Related Articles

SiliconANGLE: AMD maps $2 trillion AI market strategy to challenge Nvidia's dominance
SiliconANGLE: AMD pilots 'token routing' to slash enterprise AI costs by 43%
YouTube: NTT's LLMlet enables distributed LLM inference across browsers via WebRTC

Newest

about 21 hours ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
about 21 hours ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
about 21 hours ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
about 22 hours ago
Investing.com: TF1 Digital Revenues Jump 17% as Netflix Partnership Exceeds Growth Targets
about 22 hours ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
1 day ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
1 day ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
1 day ago
daily.dev: AVIF achieves universal browser support as Edge and Safari close gaps
2 days ago
Ealing Times: YouTube debuts UK Shopping Affiliate Programme with M&S and Currys
2 days ago
Investing.com: AMD and Cerebras debut disaggregated architecture to slash AI inference latency
2 days ago
MediaPost: Sports leagues explore non-exclusive local rights as RSN model collapses
2 days ago
YouTube: Blackmagic Design details GPU optimization protocols for DaVinci Resolve workflows
2 days ago
Startup Fortune: AI data centers threaten US grid stability and freeze cloud pipelines
2 days ago
TechRadar: OpenAI joins coalition lobbying against strict open-weight AI model regulations
2 days ago
Startup Fortune: SPAN and Nvidia board residential homes with 16-GPU Blackwell compute nodes
2 days ago
Digital Applied: Google faces €890M EU fine as Digital Markets Act enforcement accelerates
2 days ago
iZOOlogic: Ultra Clean Android App Masquerades as Utility to Host Malware-Grade Adware
2 days ago
SiliconANGLE: HPE and AMD converge supercomputing and AI via liquid-cooled GX5000
2 days ago
MarketBeat: AMD data center revenue surges 38% to $10.25B on AI demand
2 days ago
PPC Land: Acast revenue per listen jumps 26% despite flat audience growth

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.Tech Times60
  4. 4.YouTube59
  5. 5.AdExchanger57
  6. 6.TechCrunch54
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →

Newest

about 21 hours ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
about 21 hours ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
about 21 hours ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
about 22 hours ago
Investing.com: TF1 Digital Revenues Jump 17% as Netflix Partnership Exceeds Growth Targets
about 22 hours ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
1 day ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
1 day ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
1 day ago
daily.dev: AVIF achieves universal browser support as Edge and Safari close gaps
2 days ago
Ealing Times: YouTube debuts UK Shopping Affiliate Programme with M&S and Currys
2 days ago
Investing.com: AMD and Cerebras debut disaggregated architecture to slash AI inference latency
2 days ago
MediaPost: Sports leagues explore non-exclusive local rights as RSN model collapses
2 days ago
YouTube: Blackmagic Design details GPU optimization protocols for DaVinci Resolve workflows
2 days ago
Startup Fortune: AI data centers threaten US grid stability and freeze cloud pipelines
2 days ago
TechRadar: OpenAI joins coalition lobbying against strict open-weight AI model regulations
2 days ago
Startup Fortune: SPAN and Nvidia board residential homes with 16-GPU Blackwell compute nodes
2 days ago
Digital Applied: Google faces €890M EU fine as Digital Markets Act enforcement accelerates
2 days ago
iZOOlogic: Ultra Clean Android App Masquerades as Utility to Host Malware-Grade Adware
2 days ago
SiliconANGLE: HPE and AMD converge supercomputing and AI via liquid-cooled GX5000
2 days ago
MarketBeat: AMD data center revenue surges 38% to $10.25B on AI demand
2 days ago
PPC Land: Acast revenue per listen jumps 26% despite flat audience growth

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.Tech Times60
  4. 4.YouTube59
  5. 5.AdExchanger57
  6. 6.TechCrunch54
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →