inclusionAI: Ling 3.0 Flash VL
by Inclusion AItext+image+video->text1 endpoint
Overview
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual agent capabilities. Hybrid instant/reasoning model with tool calling.
Capabilities
Text generation
Image understanding
Video understanding
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
Top-P sampling
Stop sequences
Deterministic seed
Modalities
Input
TextImageVideo
Output
Text
Technical Specifications
Context Window
131.1K tokens
Max Output
32.8K tokens
Tokenizer
Other
Uptime (24h)
10000.0%
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Pricing
Per 1M tokens
Input$0.06
Output$0.18
Cache read$0.01
Added
September 10, 2026
Last synced 9/18/2026