StreamingMemeStreamingMemeBuyers Guide
AboutLeaderboardsEventsSubmit News
Subscribe

Daily Brief

The streaming industry in your inbox every morning.

Daily Brief

The streaming industry in your inbox every morning.

StreamingMemeStreamingMeme

The independent buyers guide and news aggregator for the streaming technology industry.

Explore

Buyers GuideLeaderboardsEventsSubmit News

Stay updated

Weekly digest of new companies and streaming news.

Categories

Encoding & SoftwareVideo Delivery & CDNStreaming PlatformsAI for VideoProduction HardwareBusiness NewsMonetization & Ad TechRegulatory & Policy

© 2026 StreamingMeme. All rights reserved.

AboutPrivacy PolicyTermsContact
EncodingCDNPlatformsAI & VideoHardwareBusinessAd TechPolicy
← AI for Video
AI & VideoTechnical Development

ComfyUI workflow adds lip-sync dubs in six languages

Google

This document describes a workflow for creating lip-synced video dubs and rephrases using a ComfyUI-based system, leveraging LTX-2.3-22b-IC-LoRA-LipDub models from Hugging Face. The workflow allows for translating speech or rephrasing dialogue in a video while generating new lip movements and audio to match the provided text, supporting multiple languages and emphasizing the importance of matching original dialogue length for natural-sounding output.

Key Takeaways

  • The workflow uses LTX-2.3-22b-IC-LoRA-LipDub from Hugging Face inside ComfyUI.
  • It supports both dubbing, which translates speech, and rephrasing, which changes dialogue without changing language.
  • The model generates new lip movements and audio to match the provided text.
  • The notes say the current LoRA supports only one speaker.
  • The document recommends using native script for the target language and keeping the translated line close to the original length to avoid skipped words or unnatural pacing.

Why It Matters

This is a practical pipeline for localized video edits: the workflow does not just swap audio, it also regenerates lip motion to fit the new line. That matters for dubbing and rephrasing use cases where visible mouth movement is part of the quality bar. The document also makes the operating constraints clear: one speaker only, native script, and similar line length. For teams evaluating AI video tooling, those limits are as important as the generation itself. Watch for how the same workflow behaves across the listed language prompts and whether longer or shorter translated lines degrade output.

Additional Context

The release of the LTX-2.3 LipDub workflow aligns with a broader industry shift toward diffusion-based architectures for video localization. Per lipsync.com (February 2026), diffusion models are rapidly replacing Generative Adversarial Networks (GANs) due to their superior ability to preserve subject identity and render fine mouth detail. While GAN-based tools like Wav2Lip established the category, they often suffered from training instability and visual artifacts that modern diffusion pipelines successfully resolve. This shift is fueling a market expansion for AI dubbing tools, which Market.us (May 2026) projects will grow from $2.75 billion in 2025 to nearly $19 billion by 2035. In the commercial sector, specialized platforms are increasingly moving toward all-in-one multimodal solutions. According to NYBreakings (May 2026), tools such as Magic Hour and HeyGen are winning market share by combining video generation, voice cloning, and lip-syncing into unified workflows. These platforms cater to a growing enterprise demand—highlighted by Synthesia (April 2026)—where high-fidelity lip-sync is now a standard requirement for professional business communications. The democratization of these tools allows small creators and marketing teams to bypass traditional studio costs, which RWS (January 2026) estimates can reduce localization expenses by approximately 90%. Furthermore, the technical implementation of the LTX-2.3 workflow reflects a trend toward hybrid editing. Per TodaysMagazine (May 2026), production teams are increasingly blending traditional post-production tools with custom AI nodes rather than relying on monolithic software. This modular approach, supported by open-source repositories on Hugging Face and platforms like ComfyUI, enables granular control over specific elements such as vocal timbre and frame consistency. As more companies adopt these systems, the use of proprietary watermarking and verification tools is also rising to ensure transparency in AI-modified media.


Read full article at drive.google.com

Get this in your inbox → Subscribe

Enjoy our coverage?

Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.

Add as preferred source

Related Articles

Qiang Zhang: DeltaToken cuts video tokens from 180K to under 1,000
ayushchat: Whisper runs locally on Apple Silicon with no network access
South China Morning Post: ByteDance’s Seedance 2.0 can generate feature-length films
GitHub: Lightricks LTX-2 optimization enables 4K AI video on consumer GPUs

Newest

about 10 hours ago
AOL: UK Government considers complete Freeview switch-off between 2034 and 2044
about 10 hours ago
Broadcast: Games of the Future 2026 secures global streaming and broadcast distribution
about 10 hours ago
Investing.com: Alphabet upgraded as Google Cloud revenue surges 82% on AI demand
about 10 hours ago
The Desk: Phynd launches ad-supported cloud gaming beta on LG webOS
about 11 hours ago
Kalkine Media: Adveritas hits A$16.3M recurring revenue, shifts toward cash flow breakeven
about 16 hours ago
Hyper.ai: Google DeepMind and UC Riverside launch framework to trace synthetic video
about 16 hours ago
Exame: Brazil launches TV 3.0 with 4K VVC and interactive IP layers
about 16 hours ago
Digiday: IAB Redefining Media Types Standard targets automated video ad transparency
about 16 hours ago
VentureBeat: Moonshot AI releases Kimi K3 weights with $20M revenue licensing threshold
about 16 hours ago
AdExchanger: Streaming ad tech consolidation turns independent platforms into proprietary gardens
about 16 hours ago
The Fast Mode: AMD and South Korea Partner to Build Heterogeneous Sovereign AI Infrastructure
about 16 hours ago
MediaPost: Microsoft launches Project Perception to defend programmatic supply chains from AI-driven fraud
about 16 hours ago
AdExchanger: Google mandates biometric passkeys for Ads API as AI costs reshape agency deals
about 16 hours ago
Advanced Television: Roku and Fire TV solidify gatekeeper status as OS influence grows
1 day ago
Hackernoon: Production voice pipeline solves African language latency and hallucination problems
1 day ago
Startup Fortune: Higgsfield AI integrates third-party models as revenue run rate hits $300M
1 day ago
Design & Reuse: Stricter ETSI secure boot standards mandate hardware-level chain of trust
1 day ago
UK Parliament: UK Parliament launches investigation into Ofcom's Online Safety Act enforcement
1 day ago
SVG Europe: WBD streams 600 hours of Glasgow 2026 via remote-first infrastructure
1 day ago
TradingView: Amazon settles FTC Prime suit for $2.5B amid AI-focused redesign

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group105
  2. 2.SiliconANGLE93
  3. 3.AdExchanger66
  4. 4.Tech Times65
  5. 5.YouTube62
  6. 6.TechCrunch56
  7. 7.PPC Land51
  8. 8.arXiv50
Full leaderboards →

Newest

about 10 hours ago
AOL: UK Government considers complete Freeview switch-off between 2034 and 2044
about 10 hours ago
Broadcast: Games of the Future 2026 secures global streaming and broadcast distribution
about 10 hours ago
Investing.com: Alphabet upgraded as Google Cloud revenue surges 82% on AI demand
about 10 hours ago
The Desk: Phynd launches ad-supported cloud gaming beta on LG webOS
about 11 hours ago
Kalkine Media: Adveritas hits A$16.3M recurring revenue, shifts toward cash flow breakeven
about 16 hours ago
Hyper.ai: Google DeepMind and UC Riverside launch framework to trace synthetic video
about 16 hours ago
Exame: Brazil launches TV 3.0 with 4K VVC and interactive IP layers
about 16 hours ago
Digiday: IAB Redefining Media Types Standard targets automated video ad transparency
about 16 hours ago
VentureBeat: Moonshot AI releases Kimi K3 weights with $20M revenue licensing threshold
about 16 hours ago
AdExchanger: Streaming ad tech consolidation turns independent platforms into proprietary gardens
about 16 hours ago
The Fast Mode: AMD and South Korea Partner to Build Heterogeneous Sovereign AI Infrastructure
about 16 hours ago
MediaPost: Microsoft launches Project Perception to defend programmatic supply chains from AI-driven fraud
about 16 hours ago
AdExchanger: Google mandates biometric passkeys for Ads API as AI costs reshape agency deals
about 16 hours ago
Advanced Television: Roku and Fire TV solidify gatekeeper status as OS influence grows
1 day ago
Hackernoon: Production voice pipeline solves African language latency and hallucination problems
1 day ago
Startup Fortune: Higgsfield AI integrates third-party models as revenue run rate hits $300M
1 day ago
Design & Reuse: Stricter ETSI secure boot standards mandate hardware-level chain of trust
1 day ago
UK Parliament: UK Parliament launches investigation into Ofcom's Online Safety Act enforcement
1 day ago
SVG Europe: WBD streams 600 hours of Glasgow 2026 via remote-first infrastructure
1 day ago
TradingView: Amazon settles FTC Prime suit for $2.5B amid AI-focused redesign

Upcoming Events

Jul
29–30
Buffer-Free VideoSeattle
Aug
17–20
SET EXPOSao Paulo
Sep
11–14
IBCAmsterdam
Sep
13
SportsPro Streamtime Sports LiveAmsterdam
Sep
16–18
RTC.ONKrakow
View all events →

Top Sources

  1. 1.Sports Video Group105
  2. 2.SiliconANGLE93
  3. 3.AdExchanger66
  4. 4.Tech Times65
  5. 5.YouTube62
  6. 6.TechCrunch56
  7. 7.PPC Land51
  8. 8.arXiv50
Full leaderboards →