DeepSeek V4 Pro runs at 39% off on Moyi API. .
DeepSeek: DeepSeek V4 Pro
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding, and long-horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks. Built on the same architecture as DeepSeek V4 Flash, it introduces a hybrid attention system for efficient long-context processing.
Modalities
Price (in / out)
39% off$0.4019 / $1.2058$0.66 / $1.98per 1M
Context
1M
Released
Aug 13, 2026
Providers
Moyi API routes by health and cost within the discount range and provider list you configured. When one returns an error we fall through to the next best — you don't pick a provider in the request. Configure discount range
Need a specific set of providers?
| Provider | Context | Input | Output | Cache | Latency | Throughput | Uptime | No training |
|---|---|---|---|---|---|---|---|---|
| DeepSeek | 1M | $0.66$0.4019 | $1.98$1.2058 | $0.022$0.0134 | — | — | — | No training |
| Baidu AI Cloud | 1M | $0.66$0.4019 | $1.98$1.2058 | $0.022$0.0134 | — | — | — | No training |
| Volcengine | 1M | $0.66$0.4019 | $1.98$1.2058 | $0.022$0.0134 | — | — | — | No training |
| Tencent Cloud | 1M | $0.66$0.4019 | $1.98$1.2058 | $0.022$0.0134 | — | — | — | No training |
| Third-party providers | 1M | $0.66$0.4019 | $1.98$1.2058 | $0.022$0.0134 | — | — | — | No training |
Live discount
Each provider's average effective discount, refreshed hourly.
No discount data for this model in the last 24 hours.
Performance
Latency and throughput are measured on live gateway traffic.
Latency
Time from sending the request to the first output token, broken down by provider.
LATENCY · MOYI API GATEWAY
Throughput
Tokens per second after the first output token, broken down by provider.
THROUGHPUT · MOYI API GATEWAY
Uptime
Moyi API continuously probes every provider and falls through to the next best one on error.
Uptime (24h)
—
Uptime (3d)
—
Uptime, last 3 days
—
Uptime, last 24 hours
One line per upstream provider, taking the best success rate among that provider's channels. The Moyi API line counts only the final attempt of each request — the end-to-end result after all fallbacks.
AVAILABILITY · LAST 24H · MOYI API GATEWAY
Data security
Among this model's providers, DeepSeek, Baidu AI Cloud, Volcengine, Tencent Cloud do not train on your data.
For the full data-security policy, learn more.
Need to pick your own providers? .
Quickstart
from openai import OpenAI
client = OpenAI(
base_url="https://api.moyiapi.com/v1", # the only line that changes
api_key="<MOYIAPI_API_KEY>",
)
response = client.chat.completions.create(
model="deepseek-v4-pro",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Endpoint: https://api.moyiapi.com/v1