Google Introduces Agentic Vision in Gemini 3 Flash for Active Image Understanding
By Michal Sutter
Google introduced Agentic Vision in Gemini 3 Flash, enabling the model to actively reason about images through Python code execution rather than single-pass processing. The capability delivers 5-10% quality improvement across vision benchmarks by allowing the model to iteratively inspect and analyze images.