Google

Gemini 3.5 Flash (batch)

by Googletext+image+file+audio+video->text1 endpoint

Overview

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution loops, supporting text, image, video, audio, and PDF inputs.

Defaults to medium thinking effort for faster and more cost-efficient responses, with full support for thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs.

Capabilities

Text generation
Image understanding
Audio processing
Video understanding
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
File input
Top-P sampling
Stop sequences
Deterministic seed

Modalities

Input
TextImageVideoFileAudio
Output
Text

Technical Specifications

Context Window
1.0M tokens
Max Output
65.5K tokens
Knowledge Cutoff
2025-01-01
Tokenizer
Gemini

Supported Parameters

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p

Pricing

Per 1M tokens
Input$0.75
Output$4.50
Cache read$0.07
Audio$1.50
Reasoning$4.50
Image$0.75
Added
May 19, 2026
Last synced 9/19/2026