Thinking Machines Lab, founded by ex-OpenAI executive Mira Murati, has launched multimodal AI systems named "interaction models." These systems are designed to enable real-time human-AI communication with fluid audio and video capabilities.
Generative AI video, automated editing, scene detection, AI dubbing, voice cloning, and upscaling are moving from research demos into production pipelines. Get the daily rundown on which AI video capabilities are actually shipping, and which ones are still just announcements.
Thinking Machines Lab, founded by ex-OpenAI executive Mira Murati, has launched multimodal AI systems named "interaction models." These systems are designed to enable real-time human-AI communication with fluid audio and video capabilities.
Taja has launched an AI video clipping tool designed to convert long-form recordings into multiple short-form videos optimized for social media platforms. The system automatically formats the content for various social platforms.
Google is reportedly developing a new AI video generation tool called Gemini Omni, which would enable users to create and edit realistic videos via chat. This potential tool is anticipated to be revealed at Google I/O 2026.
ShengShu Technology launched Vidu Claw, an AI CMO designed to generate complete ad campaigns from a single brief. The system utilizes Vidu Q3 in its Video Plan feature, streamlining ad creation without per-tool pricing or complex workflows.
Refiant AI, an AI/ML company specializing in model compression, has partnered with Imperial College and University College London to advance research in ultra-efficient Edge AI and model compression. The collaboration aims to develop AI/ML solutions that reduce computational and energy costs for running frontier models.
Nota, an AI model lightweighting and optimization company, announced that its "NVA" video control solution received a global Edge AI award. The solution focuses on on-device video control.
Maestro, an AI orchestration platform, aims to improve enterprise efficiency by optimizing model deployment and cost management. The platform focuses on leveraging meta models and architectural advancements, such as those seen in Jamba, to enhance AI model selection and operational efficiency.
Enterprise leaders are outlining the requirements for agentic AI deployment, emphasizing the importance of context, data speed, and governance for successful AI agent implementation in production. The discussion focuses on what is needed to move agentic AI from concept to production reality.
Experts are warning about the potential use of AI to create fake political campaign content, citing an example from March where Republicans released an AI-manipulated video of Democratic candidate James Talarico reading his past tweets. This highlights concerns about synthetic media's impact on political discourse. The article provides a single example of AI being used in political campaigns.
Thinking Machines demonstrated a preview of its near-realtime AI voice and video conversation technology, featuring new 'interaction models.' The company states that making interactivity native to the model will enhance its intelligence and collaborative effectiveness upon scaling.
This interview with Josh Mickolio from DigiKey discusses the growing shift from cloud centers to Edge Computing, specifically highlighting its impact on making Edge AI more accessible and efficient. The article focuses on the hardware revolution enabling these advancements.
BearJam, a video production company, has introduced a new "AI Video Architect" role in response to increased demand for AI-powered video production. The position focuses on leveraging artificial intelligence for video workflows. This development indicates a specialized job function emerging within the video production industry due to AI integration.
NVIDIA is making significant investments in the global AI ecosystem, focusing on developing its AI technology. The company allocated 58 trillion Korean Won (approximately $42.5 billion USD) to these efforts, impacting various applications of AI, including potential video processing advancements.
DeepFakeMaker is identified as an AI video face-swapping tool that uses realistic AI to create deepfake videos. The article describes this technology as shaping the future of digital content creation.
Snapchat has collaborated with Gucci to introduce its inaugural Sponsored AI Lens for a luxury fashion brand's advertising campaign. This initiative allows users to interact with AI-powered lenses within the Snapchat platform.
Alibaba is reportedly integrating its Qwen AI platform with its Taobao online marketplace to introduce agentic shopping experiences. This integration aims to facilitate conversational commerce and is expected to be unveiled soon.
Former Apple CEO John Sculley states that Apple must adapt its app-centric business model to an "agentic" one to compete with companies like OpenAI in the AI era. He suggests that AI agents will become more significant than traditional apps.
Adobe's stock is trading at 14.89x trailing P/E following a 61% decline from its high. The company offers AI capabilities through its Firefly product and integrates Anthropic's Claude AI.
OpenAI has released three new real-time voice models, GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper, for its API. These models incorporate "GPT-5-class reasoning" capabilities.
Dan Hartley and Chris Bird have launched CineMe, an AI development tool for the film, TV, and gaming industries. The tool aims to leverage AI to assist with various aspects of content creation, from concept to distribution.
Christopher Patrick of Qualcomm joined Tech360 to discuss the future of generative AI, focusing on specific applications like on-device AI and Snapdragon's AI vision.
Preseem, a Quality of Experience (QoE) platform for ISPs, and QueSee AI, a customer experience and retention intelligence platform, have integrated their services. This integration allows network QoE data to be combined with customer experience analytics, providing ISPs with insights into customer churn risk and service issues. The combined solution aims to help ISPs proactively address customer concerns by correlating network performance with customer sentiment.
SoundHound's OASYS platform, designed to auto-build and improve AI agents across various channels, is highlighted for its potential to redefine agentic AI. This development coincides with LivePerson's expanded reach and SoundHound's reported 52% revenue jump to $44.2 million.
Nine out of ten autonomous AI agents deployed in production environments are vulnerable to a specific class of attack. This vulnerability affects real-world deployments, indicating a gap in standard safety testing. The article details the prevalence of these failures.
This article discusses the emerging reality of rogue artificial intelligence (AI) within corporate systems, describing the challenge of managing "AI agents" that operate autonomously. It notes that the once-theoretical concept of AI agents acting independently is now becoming a practical concern for companies.
ImageKit launched DAM Agent, a new native AI assistant designed for Digital Asset Management (DAM) platforms. This tool uses conversational AI to classify, analyze, and manage digital media assets. The release aims to enhance efficiency in asset discovery and organization.
SipRadius launched the Ferox Films AI Content Creation Platform at BroadcastAsia, a unified interface designed to streamline access to various AI tools for video content creation. The platform aims to simplify the use of AI in production workflows.
Encompass Digital Media and VideoMagic International launched 'Altitude Intelligence' at NAB 2026. This AI-driven platform is designed for the streaming industry.
AWS announced the general availability of Claude Platform on AWS, providing customers direct access to Anthropic's Claude AI models through their AWS accounts. This service integrates Anthropic's native platform with AWS's infrastructure.
Netflix is reportedly testing an enhanced AI-powered voice search feature for its streaming platform. The article describes this as an advancement for streaming UI, suggesting its potential integration could improve user experience on devices like the Apple TV 4K.
Google is developing a new video generation model named "Omni" for its Gemini AI platform. Early demonstrations suggest the model can produce video at a qualitative level that is described as impressive.
OpenAI has introduced GPT-Realtime-2, an expansion of its Realtime API. This update includes new translation and transcription models designed to enhance the speed and capability of AI voice agents.
Former Prime Video UK head Chris Bird has launched two AI-powered businesses aimed at filmmakers: HawksHead AI, a predictive data analytics platform, and CineMe, an AI visual development tool. The platforms utilize artificial intelligence to support content creation and decision-making for film production.
YouTube is expanding its automated likeness detection technology to identify AI-generated content featuring celebrities and individuals represented by talent agencies. This initiative aims to address misleading AI-generated video content on its platform.
AWS published a blog post detailing how to build web search-enabled agents using Strands and Exa within the Amazon Bedrock AgentCore framework. The article focuses on leveraging AI orchestration for enhanced agent capabilities.
Zero Latency has implemented Red Hat AI Factory, which incorporates NVIDIA technology, to serve as the Kubernetes foundation for its distributed AI inference network, Neocloud. This deployment aims to enable AI at scale for distributed applications managed by Zero Latency.
IBM's Think 2026 event highlighted advancements in 'agentic AI,' focusing on managing its speed, scale, and sprawl. The event featured keynotes, spotlights, announcements, and demos relevant to AI for business and IT leaders.
Korean public institutions are reportedly selecting Nvidia GPUs for artificial intelligence-related applications, despite a governmental 'K-Nvidia' policy aimed at supporting domestic Neutral Processing Unit (NPU) firms. This indicates a preference for established foreign technology over localized solutions in the public sector.
Apple's research interest in AI models extends to spatial understanding and sign language annotation, indicating continued development in AI applications beyond current product speculation. The company's studies explore using large language models (LLMs) for understanding spatial data and for creating sign language annotations.
Former Prime Video UK chief Chris Bird has launched two artificial intelligence ventures focusing on supporting independent content creators. One of these ventures is in partnership with documentary director Dan Hartley.
This article discusses the accelerating operational deployment of AI, including agentic and generative AI, for digital video. It notes that social video is notably outpacing CTV in the adoption of these AI technologies.
Meta is developing an agentic ecosystem, codenamed "Hatch," which involves deploying AI agents across its social networks. This initiative specifically aims to integrate agent-assisted purchasing features into Instagram.
The article discusses the evolving landscape of video training data and multimodal foundation models in 2026, shifting towards highly curated and licensed data. This trend enables the development of advanced AI applications for video, moving beyond simple quantitative data acquisition.
Genvid has integrated Alibaba's "Happy Horse" AI video generation model into its platform, enabling users to create up to 15 seconds of 1080p video with synchronized audio-visual output from text or image prompts. The update also includes enhancements to the keyframe editor and expanded ComfyUI integration with C2PA provenance.
Google is reportedly developing a new video generation AI model named Gemini Omni. This follows the company's previous advancement in the video generation space with its Veo model.
Bell Media announced plans to digitize its entire media archive and make it available on YouTube. This initiative will utilize Google's Gemini AI for the digitization process.
A study conducted by UNSW Sydney and QUT, analyzing over 435,000 Facebook ads from 891 Australians, found that large language models (LLMs) can infer user demographics such as gender and age solely from advertising activity patterns.
The WNBA has entered into a multi-year partnership with Amazon Web Services (AWS), designating AWS as its Official Cloud and Cloud AI Partner. This collaboration indicates a strategic move by the WNBA to leverage cloud and AI technologies, likely for enhanced operations or fan engagement.
TUTT conducted a 168-hour field test comparing the TUTT GS10, Ray-Ban Meta (Gen 2), and Oakley Meta Vanguard AI glasses in Toronto, evaluating their AI processing architectures, imaging systems, real-time translation capabilities, and battery performance. The test detailed performance differences in cloud-hybrid vs. edge AI processing, camera parallax, acoustic engineering, and battery endurance under varying conditions. The article concludes with a ranking based on market role, identifying each device's strengths and target users.
Zyphra and AMD have jointly launched an open-source AI platform, which utilizes MI355X GPUs for its power. The platform is designed to compete with DeepSeek and has plans for future expansion to MI450 GPUs and beyond.