Introducing agentic video understanding with Gemini
Article image or reusable cover for Google DeepMind
Google DeepMind is launching a new video analysis feature for Gemini models (3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite) that reduces energy consumption by up to 88 percent and costs by up to 66 percent while increasing accuracy by up to 7 percent.
Instead of analyzing video at a fixed rate, this "agentic" feature performs intelligent searches through video files — it can search for specific moments, detect anomalies, or count objects without loading the entire video. The feature is available now via Gemini API and will soon be available in the Gemini app for regular users.
agentic video understanding pairs the model's core reasoning with native video tools to dynamically search, scan, and inspect target video segments across visual frames, audio, and transcripts.
Read the full story at Google DeepMind →
Vibekollen prepared this summary with AI from the original publication. The content belongs to Google DeepMind.