StreamingMemeStreamingMemeBuyers Guide
AboutLeaderboardsEventsSubmit News
Subscribe

Daily Brief

The streaming industry in your inbox every morning.

Daily Brief

The streaming industry in your inbox every morning.

StreamingMemeStreamingMeme

The independent buyers guide and news aggregator for the streaming technology industry.

Explore

Buyers GuideLeaderboardsEventsSubmit News

Stay updated

Weekly digest of new companies and streaming news.

Categories

Encoding & SoftwareVideo Delivery & CDNStreaming PlatformsAI for VideoProduction HardwareBusiness NewsMonetization & Ad TechRegulatory & Policy

© 2026 StreamingMeme. All rights reserved.

AboutPrivacy PolicyTermsContact
EncodingCDNPlatformsAI & VideoHardwareBusinessAd TechPolicy
← Production Hardware
HardwareTechnical DevelopmentJuly 7, 2026

DeepSeek joins OpenAI in move to custom inference silicon

DeepSeek joins OpenAI in move to custom inference silicon
SiliconANGLE

Chinese AI startup DeepSeek is reportedly developing in-house custom inference chips to reduce dependency on external suppliers like Nvidia and Huawei. This strategic move aims to optimize commercialization costs and improve control over the computational stack as demand for high-performance AI inference grows.

Key Takeaways

  • Project focused specifically on inference chips to lower recurring commercialization costs for models like R1 and V4.
  • Company recently raised $7.4 billion from Chinese investors, providing capital to fund expensive silicon R&D.
  • Hardware strategy seeks independence from Huawei, which currently controls roughly 50% of China's $50 billion AI chip market.
  • Move follows OpenAI's June 2024 launch of Jalapeño, its first custom inference chip co-designed with Broadcom.

Why It Matters

Vertical integration is becoming the standard for frontier AI labs seeking to escape the margin squeeze of third-party hardware. For DeepSeek, building custom silicon is a survival necessity due to tightening U.S. export controls on Nvidia's H800 and Rubin architectures. Within the Chinese market, this project directly challenges Huawei's dominance as the primary local alternative. Success would allow DeepSeek to optimize performance-per-watt specifically for its reasoning models, theoretically offering a better price-to-performance ratio than rivals using generic hardware. Watch for a potential partnership with a domestic foundry like SMIC as the project moves toward a tape-out date.

Additional Context

The strategic pivot into hardware reflects a broader industry trend of large-scale model developers seeking to optimize 'the cost of intelligence' at the silicon layer. Per OpenAI and Broadcom in June 2026, their first custom Intelligence Processor, Jalapeño, achieved a 9-month design-to-production cycle using AI-assisted engineering. This chip is claimed to offer roughly 50% lower inference costs per token compared to general-purpose GPUs. In the same window, Microsoft and other hyperscalers reportedly committed to a 10-gigawatt infrastructure roadmap specifically built around these custom accelerators to support the next generation of agentic AI workflows. In China, the push for self-reliance has accelerated as the U.S. Department of Commerce moved in June 2026 to close loopholes regarding the export of advanced Blackwell processors to overseas subsidiaries of Chinese firms, per Reuters. This regulatory pressure has forced domestic leaders to lean on local infrastructure; for instance, DeepSeek’s V4 series was specifically optimized for Huawei’s Ascend 950PR and 950DT chips. However, per Asia Tech Review in July 2026, leading startups like Zhipu AI are joining DeepSeek in pursuing bespoke processors to avoid becoming overly dependent on a single domestic vendor like Huawei. Huawei remains the incumbent leader, confirmng in June 2026 that its new Ascend 950DT chip—featuring 4 TB/s memory bandwidth—will debut on its cloud ecosystem in August 2026 to support massive LLM clusters. Despite this domestic progress, custom chip ventures from startups like DeepSeek face significant manufacturing hurdles, as U.S. restrictions continue to bar Chinese designers from reaching the most advanced sub-3nm global foundries required for efficiency parity with Western peers.


Read full article at siliconangle.com

Get this in your inbox → Subscribe

Enjoy our coverage?

Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.

Add as preferred source

Related Articles

Radio World: Orban Labs transitions to Linux-based processing using Raspberry Pi clusters
YouTube: Blackmagic Design details GPU optimization protocols for DaVinci Resolve workflows
Radio World: Broadcast Consultant Urges Simpler Air Chains and Shift to Linear Audio

Newest

1 day ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
1 day ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
1 day ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
1 day ago
Investing.com: TF1 Digital Revenues Jump 17% as Netflix Partnership Exceeds Growth Targets
1 day ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
1 day ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
1 day ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
1 day ago
daily.dev: AVIF achieves universal browser support as Edge and Safari close gaps
2 days ago
Ealing Times: YouTube debuts UK Shopping Affiliate Programme with M&S and Currys
2 days ago
Investing.com: AMD and Cerebras debut disaggregated architecture to slash AI inference latency
2 days ago
MediaPost: Sports leagues explore non-exclusive local rights as RSN model collapses
2 days ago
YouTube: Blackmagic Design details GPU optimization protocols for DaVinci Resolve workflows
2 days ago
Startup Fortune: AI data centers threaten US grid stability and freeze cloud pipelines
2 days ago
TechRadar: OpenAI joins coalition lobbying against strict open-weight AI model regulations
2 days ago
Startup Fortune: SPAN and Nvidia board residential homes with 16-GPU Blackwell compute nodes
2 days ago
Digital Applied: Google faces €890M EU fine as Digital Markets Act enforcement accelerates
2 days ago
iZOOlogic: Ultra Clean Android App Masquerades as Utility to Host Malware-Grade Adware
2 days ago
SiliconANGLE: HPE and AMD converge supercomputing and AI via liquid-cooled GX5000
2 days ago
MarketBeat: AMD data center revenue surges 38% to $10.25B on AI demand
2 days ago
PPC Land: Acast revenue per listen jumps 26% despite flat audience growth

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.Tech Times60
  4. 4.YouTube59
  5. 5.AdExchanger57
  6. 6.TechCrunch54
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →

Newest

1 day ago
Barchart: Cerebras and AMD partner on low-latency AI inference architecture
1 day ago
Light Reading: Charter sidesteps Starlink partnership rumors as Q2 broadband losses widen
1 day ago
GuruFocus: Fastly joins Experian to secure autonomous commerce at the edge
1 day ago
Investing.com: TF1 Digital Revenues Jump 17% as Netflix Partnership Exceeds Growth Targets
1 day ago
The BIG Newsletter: Nexstar and TEGNA Accused of Violating Judicial Order in $6.2 Billion Merger
1 day ago
Vocal: TeqBlaze challenges Epom with modular full-stack white-label ad tech suite
1 day ago
Audio Chocolate: Merging Technologies debuts Anubis Premium SPS for mission-critical broadcast audio
1 day ago
daily.dev: AVIF achieves universal browser support as Edge and Safari close gaps
2 days ago
Ealing Times: YouTube debuts UK Shopping Affiliate Programme with M&S and Currys
2 days ago
Investing.com: AMD and Cerebras debut disaggregated architecture to slash AI inference latency
2 days ago
MediaPost: Sports leagues explore non-exclusive local rights as RSN model collapses
2 days ago
YouTube: Blackmagic Design details GPU optimization protocols for DaVinci Resolve workflows
2 days ago
Startup Fortune: AI data centers threaten US grid stability and freeze cloud pipelines
2 days ago
TechRadar: OpenAI joins coalition lobbying against strict open-weight AI model regulations
2 days ago
Startup Fortune: SPAN and Nvidia board residential homes with 16-GPU Blackwell compute nodes
2 days ago
Digital Applied: Google faces €890M EU fine as Digital Markets Act enforcement accelerates
2 days ago
iZOOlogic: Ultra Clean Android App Masquerades as Utility to Host Malware-Grade Adware
2 days ago
SiliconANGLE: HPE and AMD converge supercomputing and AI via liquid-cooled GX5000
2 days ago
MarketBeat: AMD data center revenue surges 38% to $10.25B on AI demand
2 days ago
PPC Land: Acast revenue per listen jumps 26% despite flat audience growth

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group104
  2. 2.SiliconANGLE91
  3. 3.Tech Times60
  4. 4.YouTube59
  5. 5.AdExchanger57
  6. 6.TechCrunch54
  7. 7.arXiv50
  8. 8.PPC Land48
Full leaderboards →