Google has introduced agentic video understanding, a new approach that lets Gemini decide which parts of a video deserve closer inspection instead of processing the entire timeline at a fixed rate.

Google says the approach can cut token consumption by up to 88%, reduce analysis costs by up to 66% and improve accuracy by up to 7% on standard video benchmarks, potentially lowering the cost of applications that analyze long recordings.

Traditional video processing, which Google calls static processing, typically samples one frame per second and sends those frames to the model. Developers can change the frame rate, but the model still has to…

Read the full article at TECHREPUBLIC.COM