NVIDIA's synthetic-video detector spots AI-generated content with 92%+ accuracy in 22ms

NVIDIA introduced a synthetic-video detection NIM microservice that identifies AI-generated video and fake news content with a reported internal AUC of 0.9614 and 94.5% accuracy. Critically for real-world deployment, it processes 1080p footage in as little as 22 milliseconds on RTX systems and roughly 30ms on L40 GPUs — fast enough for near-real-time moderation pipelines at scale.
The tool addresses a rapidly worsening problem: as generative video models (and the week's flurry of open image/video models) proliferate, distinguishing authentic footage from synthetic becomes essential for newsrooms, platforms, and trust-and-safety teams. Packaging it as a NIM microservice makes it drop-in deployable within NVIDIA's inference stack.
Alongside the detector, NVIDIA and Hugging Face expanded the open-source LeRobot platform with new robotics AI tools, reinforcing NVIDIA's physical-AI strategy that also produced Cosmos 3 Edge this week. Separately, inference startup Infinity raised $15 million from Touring Capital with backing from OpenAI and Anthropic researchers — a signal that the inference-efficiency race (echoing Google's Frozen v2 and DeepSeek's chip plans) is drawing serious investor and researcher attention.
Skeptics note that detector accuracy figures are vendor-reported and internal, and that synthetic-media detection is an arms race where generators quickly adapt to evade classifiers — today's 94.5% can erode fast. Real-world false-positive rates on diverse content also tend to lag lab numbers. Readers should watch for independent evaluation, adoption by major platforms, and how quickly newer generative models defeat the detector.