Gemini 3.5 Flash (batch)
by Googletext+image+file+audio+video->text1 endpoint
Overview
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution loops, supporting text, image, video, audio, and PDF inputs.
Defaults to medium thinking effort for faster and more cost-efficient responses, with full support for thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs.
Capabilities
Text generation
Image understanding
Audio processing
Video understanding
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
File input
Top-P sampling
Stop sequences
Deterministic seed
Modalities
Input
TextImageVideoFileAudio
Output
Text
Technical Specifications
Context Window
1.0M tokens
Max Output
65.5K tokens
Knowledge Cutoff
2025-01-01
Tokenizer
Gemini
Supported Parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Pricing
Per 1M tokens
Input$0.75
Output$4.50
Cache read$0.07
Audio$1.50
Reasoning$4.50
Image$0.75
Added
May 19, 2026
Last synced 9/19/2026