Google

Gemini 3 Flash Preview (batch)

by Googletext+image+file+audio+video->text1 endpoint

Overview

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive development, long running agent loops, and collaborative coding tasks. Compared to Gemini 2.5 Flash, it provides broad quality improvements across reasoning, multimodal understanding, and reliability.

The model supports a 1M token context window and multimodal inputs including text, images, audio, video, and PDFs, with text output. It includes configurable reasoning via thinking levels (minimal, low, medium, high), structured output, tool use, and automatic context caching. Gemini 3 Flash Preview is optimized for users who want strong reasoning and agentic behavior without the cost or latency of full scale frontier models.

Capabilities

Text generation
Image understanding
Audio processing
Video understanding
Tool use / Function calling
Structured output
Extended reasoning
Long context
File input
Top-P sampling
Stop sequences
Deterministic seed

Modalities

Input
TextImageFileAudioVideo
Output
Text

Technical Specifications

Context Window
1.0M tokens
Max Output
65.5K tokens
Tokenizer
Gemini

Supported Parameters

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p

Pricing

Per 1M tokens
Input$0.25
Output$1.50
Audio$0.50
Reasoning$1.50
Image$0.25
Added
December 17, 2025
Last synced 9/19/2026