MODELS
by Google DeepMind
Fast Gemini 3.5 model for reasoning, coding, long-context, multimodal, and agentic tool-use workflows.
Gemini 3.5 Flash is a Google DeepMind Gemini model described in official sources as a fast model for reasoning, coding, long-context, multimodal, and agentic tool-use workflows. The Gemini API model page lists model code gemini-3.5-flash, supports text, image, video, audio, and PDF inputs with text output, and documents a 1,048,576-token input limit and 65,536-token output limit. Official docs also list capabilities such as function calling, structured outputs, grounding, thinking, URL context, code execution, and computer use preview.
Use for Gemini API applications that need a fast Gemini 3.5 model with reasoning and coding support. Use for workflows requiring long-context input with the documented 1 048 576-token input limit. Use for multimodal input workflows involving text image video audio or PDFs with text output. Use when an application needs supported Gemini API capabilities such as function calling structured outputs grounding URL context code execution or computer use preview.
Last updated Jul 31, 2026