Gemini 3.5 Flash runs at 76% off on Moyi API. .
Google: Gemini 3.5 Flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution loops, supporting text, image, video, audio, and PDF inputs. Defaults to medium thinking effort for faster and more cost-efficient responses, with full support for thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs.
Modalities
Price (in / out)
76% off$0.354 / $2.124$1.5 / $9per 1M
Context
1M
Released
May 19, 2026
Providers
Moyi API routes by health and cost within the discount range and provider list you configured. When one returns an error we fall through to the next best — you don't pick a provider in the request. Configure discount range
Need a specific set of providers?
| Provider | Context | Input | Output | Cache | Latency | Throughput | Uptime | No training |
|---|---|---|---|---|---|---|---|---|
| Google AI | 1M | $1.5$0.354 | $9$2.124 | $0.15$0.0354 | — | — | — | No training |
| Google Vertex | 1M | $1.5$0.354 | $9$2.124 | $0.15$0.0354 | — | — | — | No training |
| Third-party providers | 1M | $1.5$0.354 | $9$2.124 | $0.15$0.0354 | — | — | — | No training |
Live discount
Each provider's average effective discount, refreshed hourly.
No discount data for this model in the last 24 hours.
Performance
Latency and throughput are measured on live gateway traffic.
Latency
Time from sending the request to the first output token, broken down by provider.
LATENCY · MOYI API GATEWAY
Throughput
Tokens per second after the first output token, broken down by provider.
THROUGHPUT · MOYI API GATEWAY
Uptime
Moyi API continuously probes every provider and falls through to the next best one on error.
Uptime (24h)
—
Uptime (3d)
—
Uptime, last 3 days
—
Uptime, last 24 hours
One line per upstream provider, taking the best success rate among that provider's channels. The Moyi API line counts only the final attempt of each request — the end-to-end result after all fallbacks.
AVAILABILITY · LAST 24H · MOYI API GATEWAY
Data security
Among this model's providers, Google AI, Google Vertex do not train on your data.
For the full data-security policy, learn more.
Need to pick your own providers? .
Quickstart
from google import genai
from google.genai import types
client = genai.Client(
api_key="<MOYIAPI_API_KEY>",
http_options=types.HttpOptions(base_url="https://api.moyiapi.com"),
)
response = client.models.generate_content(
model="gemini-3.5-flash",
contents="Hello",
)
print(response.text)Endpoint: https://api.moyiapi.com