Gemini 3.8 Flash released, described as most intelligent Gemini model to date
Google DeepMind launched agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite on September 1, letting the models dynamically choose what to watch, at what speed, and through which modality rather than processing at a fixed frame rate. The company says the feature drops token consumption by up to 88% and costs by up to 66%, while lifting accuracy by up to 7% on standard benchmarks, and places Gemini 3.7 Flash at the accuracy-to-cost Pareto frontier among tested models. New capabilities include sub-second moment retrieval, anomaly detection, and long-form needle-in-a-haystack search; Google DeepMind says the feature will power YouTube's 'Ask YouTube' and roll out to Gemini app users across Flash and Flash-Lite models.