Gemini Introduces Agentic Video Understanding for Enhanced Analysis
Google is launching agentic video understanding across its Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite models today. This capability enhances video analysis by dynamically searching and inspecting video segments, reducing token consumption by up to 88% and costs by 66% while improving accuracy by 7%. It enables advanced features like sub-second moment retrieval and precise anomaly detection for long-form content. Developers can access this feature immediately via the Gemini API, with wider integration into Google products coming soon.
Features (1) ›
- Agentic Video Understanding for Gemini Models
Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite now support agentic video understanding, enhancing video analysis by dynamically searching and inspecting target segments across visual frames, audio, and transcripts. This reduces analysis costs by up to 66% and token consumption by 88% while improving accuracy by 7%, especially for long-form videos. It unlocks capabilities like sub-second moment retrieval, accurate anomaly detection, and complex queries, and is available today via the Gemini API in Google AI Studio and the Gemini Enterprise Agent Platform.
https://blog.google/innovation-and-ai/models-and-research/gemini-models/introducing-agentic-video-in-gemini/
Related releases
- Python Generative AI SDK v2.21.0 Adds Streaming Downloads, WEBM Audio, Video Understanding Gemini Python SDK Releases ·
- Gemini model gemini-omni-flash-preview reaches end of life in 30 days endoflife.date ·
- Gemini model gemini-robotics-er-1.6-preview has reached end of life endoflife.date ·
- Gemini model gemini-omni-flash-preview reaches end of life in 33 days endoflife.date ·
- Gemini Omni 1.1 Flash Now Accessible via APIs, Offering Enhanced Control Gemini Blog ·
- Google showcases 7 ways Gemini enhances student productivity in Workspace Gemini Blog ·