DeepSeek: DeepSeek V4 Flash Vision Exp
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents, reasoning, and world knowledge. It is a sparse mixture-of-experts model with 13B active parameters out of 284B total. It is suited for document and chart understanding, visual question answering, and multimodal agent workflows that interleave text and images.
Modalities
Price (in / out)
$0.22 / $0.66per 1M
Context
1M
Released
Aug 21, 2026
Providers
Moyi API routes by health and cost within the discount range and provider list you configured. When one returns an error we fall through to the next best — you don't pick a provider in the request. Configure discount range
Need a specific set of providers?
| Provider | Context | Input | Output | Cache | Latency | Throughput | Uptime | No training |
|---|---|---|---|---|---|---|---|---|
| DeepSeek | 1M | $0.22 | $0.66 | $0.007 | — | — | — | No training |
| Baidu AI Cloud | 1M | $0.22 | $0.66 | $0.007 | — | — | — | No training |
| Volcengine | 1M | $0.22 | $0.66 | $0.007 | — | — | — | No training |
| Tencent Cloud | 1M | $0.22 | $0.66 | $0.007 | — | — | — | No training |
| Third-party providers | 1M | $0.22 | $0.66 | $0.007 | — | — | — | No training |
Live discount
Each provider's average effective discount, refreshed hourly.
No discount data for this model in the last 24 hours.
Performance
Latency and throughput are measured on live gateway traffic.
Latency
Time from sending the request to the first output token, broken down by provider.
LATENCY · MOYI API GATEWAY
Throughput
Tokens per second after the first output token, broken down by provider.
THROUGHPUT · MOYI API GATEWAY
Uptime
Moyi API continuously probes every provider and falls through to the next best one on error.
Uptime (24h)
—
Uptime (3d)
—
Uptime, last 3 days
—
Uptime, last 24 hours
One line per upstream provider, taking the best success rate among that provider's channels. The Moyi API line counts only the final attempt of each request — the end-to-end result after all fallbacks.
AVAILABILITY · LAST 24H · MOYI API GATEWAY
Data security
Among this model's providers, DeepSeek, Baidu AI Cloud, Volcengine, Tencent Cloud do not train on your data.
For the full data-security policy, learn more.
Need to pick your own providers? .
Quickstart
from openai import OpenAI
client = OpenAI(
base_url="https://api.moyiapi.com/v1", # the only line that changes
api_key="<MOYIAPI_API_KEY>",
)
response = client.chat.completions.create(
model="deepseek-v4-flash-vision-exp",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Endpoint: https://api.moyiapi.com/v1