Gemini now analyzes your videos more accurately, and at a lower cost Google has introduced agentic video understanding in its latest Gemini models—Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite—enabling more accurate and cost-efficient video analysis, with token usage reduced by up to 88% and accuracy improved by up to 7%. The feature is currently available via the Gemini API in Google AI Studio and the Gemini Enterprise Agent Platform, with plans to roll out to the Gemini app and power YouTube's 'Ask YouTube' feature. Affiliate links on Android Authority may earn us a commission. Learn more. https://www.androidauthority.com/external-links/ Gemini now analyzes your videos more accurately, and at a lower cost Sep 2, 2026 — 1:28 AM ET - Google has introduced agentic video understanding in Gemini. - Gemini can now easily analyze long videos with sub-second accuracy. - It can also track physical movement and count distinct objects in videos. Google has been steadily improving Gemini https://www.androidauthority.com/gemini-tips-and-tricks-android-3644941/ to make it more useful. We’ve already spotted the company working on new features https://www.androidauthority.com/gemini-spark-chat-assign-apk-teardown-3705365/ for Gemini on mobile, while also releasing new Gemini models such as Gemini 3.7 Flash https://www.androidauthority.com/gemini-3-7-flash-debut-3698440/ . However, Google is now giving Gemini the ability to analyze videos in a more cost-effective and efficient manner. Google announced https://blog.google/innovation-and-ai/models-and-research/gemini-models/introducing-agentic-video-in-gemini/ today that it’s bringing agentic video understanding to the latest Gemini models: Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite. This new capability allows users to upload videos to Gemini and have the AI analyze them. It also reduces token usage by up to 88% and offers up to 7% better accuracy. Users have been able to upload videos to Gemini for a while now, but the AI could only perform what Google calls “static” processing on them. This meant it would split the video into individual frames, resulting in slower performance and higher costs. With agentic video understanding, Gemini can now decide what to watch and at what speed. It can also choose between frames, audio, and the transcript to analyze videos more efficiently. This means Gemini can pinpoint split-second changes, answer complex questions across multi-hour videos, inspect videos for visual artifacts, and even count and track physical movements and objects. Right now, agentic video understanding is only available via the Gemini API in Google AI Studio and the Gemini Enterprise Agent Platform. However, the company also said the feature will roll out to the Gemini app soon. Google will also soon start using agentic video understanding to power YouTube’s “Ask YouTube” feature, which could make it easier for creators to analyze their videos with Gemini. Thank you for being part of our community. Read our Comment Policy https://www.androidauthority.com/android-authority-comment-policy/ before posting.