A Gartner survey of 297 cybersecurity leaders reveals that 41% of CISOs encountered deepfake-based social engineering in audio calls over the past year. The report advises organizations to shift from static training to adaptive verification protocols and hardened identity controls to mitigate AI-driven impersonation risks.
The high prevalence of deepfake social engineering signals a breakdown in traditional identity verification for streaming and media organizations. As attackers use synthetic media to impersonate executives in audio and video calls, the industry must move beyond static security training toward hardened identity controls and phishing-resistant authentication. For streaming platforms managing high-value content rights and sensitive user data, this shift requires integrating threat detection with account recovery and financial transaction workflows. The ecosystem must now treat every multimodal communication as a potential risk regardless of the channel's perceived intimacy. Watch for organizations to implement mandatory out-of-band verification for any high-risk administrative or financial request.
Gartner's deepfake findings arrive as video-platform vendors race to build detection and verification layers directly into their stacks. In May 2026, Bitmovin's annual Video Developer Report found that 98% of 486 respondents now use AI or ML somewhere in their video workflows, with 46% employing AI tools daily. The same report noted that visual quality and optimisation (30%) and tagging and categorisation (28%) are among the top AI applications, categories where synthetic media detection is a natural extension. Bitmovin CEO Stefan Lederer observed that the last three months alone reshaped what video developers can build, underscoring how quickly the tooling landscape is shifting beneath security teams evaluating deepfake countermeasures.
On the business and deployment side, Bitmovin has been expanding its enterprise footprint in ways that intersect with content authenticity concerns. In May 2026, Bitmovin announced that MUBI selected its VOD Encoder to replace a legacy on-premises stack, supporting 3-pass encoding, UHD, and a multi-codec strategy spanning AVC, HEVC, and AV1. The deal, signed through AWS Marketplace, illustrates how premium content platforms are consolidating encoding infrastructure with vendors that carry SOC 2 Type II and ISO 27001 certifications, compliance frameworks that increasingly matter when organizations must demonstrate chain-of-custody controls against synthetic media threats.
Competing video-infrastructure vendors are also layering AI capabilities that touch on content verification. Mux in 2026 launched Mux Robots, a first-party API that runs video analysis jobs natively inside its platform, building on the open-source @mux/ai toolkit released in December 2025. The product handles moderation, summarisation, and Q&A workflows with automatic provider selection, and Mux recently introduced Robots Directives for multi-step orchestration. While Mux Robots targets content moderation rather than deepfake detection per se, the architectural pattern of embedding AI analysis directly alongside video assets is the same approach security vendors are adopting to flag synthetic audio and video in real time. For streaming platforms evaluating deepfake social engineering defenses, the question is whether detection belongs in the encoding and delivery pipeline or in a separate identity-verification layer, and both Bitmovin and Mux are positioning their platforms to host that workload.
A recent Gartner survey of 297 CISOs found that 41% encountered deepfake social engineering via audio calls over the past year. This trend signals a breakdown in traditional identity verification, forcing organizations to move beyond static training toward mandatory out-of-band verification and phishing-resistant authentication for all high-value administrative and financial requests.
According to the Gartner survey, 41% of the 297 surveyed CISOs reported experiencing deepfake social engineering incidents during audio calls.
The Gartner survey found that 36% of CISOs encountered deepfake social engineering incidents during video calls.
Gartner analyst Craig Porter recommends that organizations move away from 'spot the fake' training and instead implement mandatory verification protocols for all consequential requests.
Yes, vendors like Bitmovin and Mux are embedding AI analysis directly into video pipelines, which provides an architectural pattern that security teams can use to flag synthetic audio and video in real time.
Add StreamingMeme as a preferred source on Google to see more of our streaming news at the top of your Search results.
Add as preferred source