← all models
Umans DeepSeek V4 Flash Vision (lab) Experimental
umans-deepseek-v4-flash-0731-vision-lab · DeepSeek-V4-Flash-Vision · DeepSeek
In testing
throughput · p50 · no requests
TTFT · p50 · no requests
uptime · 24h

DeepSeek V4 Flash Vision as a Labs experiment, open for a short test window: temporary, not a permanent id. Our own V4 Flash with vision added: the same performance and speed as DeepSeek V4 Flash (284B total, 13B active, 1M-token context), plus the ability to read images. Reasoning has four modes: non-think (none), think low (low, the default), think high (high) and think max (max). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and low availability, and it is served on our own GPU infrastructure, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again.

no uptime history yet
90 days agotoday
Context
1049K
Max output
393K
Recommended
393K
Vision
Yes
Tools
Yes
Reasoning
Toggle · none/low/high/max
Trends

Speed over the last 90 days

daily medians · dashed line = target
throughput p50 · output tokens per second, higher is better
gathering data
90 days agotoday
TTFT p50 · time to first token, lower is better
gathering data
90 days agotoday
Changelog

Events for Umans DeepSeek V4 Flash Vision (lab)

incl. gateway-wide announcements
Aug 112026
New Labs experiment: Umans DeepSeek V4 Flash Vision Testing
The text V4 Flash lab window ended and a new vision lab opened on umans-deepseek-v4-flash-0731-vision-lab: our own V4 Flash with vision added, the same performance and speed as the text V4 Flash plus the ability to read images. Free and seat-gated while the experiment runs, served at limited capacity and low availability: it is a lab, so expect it to be flaky and to go down under load - crash it, give it a moment, and try again.