DeepSeek

DeepSeek V4 Flash Vision Exp

by DeepSeektext+image->text5 endpoints

Overview

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents, reasoning, and world knowledge. It is a sparse mixture-of-experts model with 13B active parameters out of 284B total.

It is suited for document and chart understanding, visual question answering, and multimodal agent workflows that interleave text and images.

Capabilities

Text generation
Image understanding
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
Top-P sampling
Stop sequences
Deterministic seed
Log probabilities

Modalities

Input
TextImage
Output
Text

Technical Specifications

Context Window
1.0M tokens
Max Output
262.1K tokens
Tokenizer
DeepSeek
Uptime (24h)
9986.7%
License
unknownOpen Weights
Country
🇨🇳CN
Release Date
August 21, 20261mo

Supported Parameters

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Pricing

Per 1M tokens
Input$0.22
Output$0.65
Cache read$0.0069
Added
August 21, 2026
Last synced 9/20/2026