Gemini 3.5 Flash
ZDRLatest Gemini Flash - frontier performance, standard tier
Providers
Different companies host the same model. DOS.AI routes your request to one of them based on the routing mode you pick — Balanced (price + speed), Nitro (fastest), or Exacto (highest tool-calling accuracy).
| Provider | Input /1M | Output /1M | Latency | Uptime |
|---|---|---|---|---|
Google | $1.50 | $9.00 | - | - |
Pricing
Rates from the DOS.AI catalog for the billing units shown. Caching and provider discounts mean the price actually paid is often below the listed one.
Provider↕ | Effective in /M↕ | Effective out /M↕ | Cache read /M↕ | Latency↕ | Uptime↕ | Savings↕ |
|---|---|---|---|---|---|---|
Google | $1.50 | $9.00 | - | - | - | - |
Observed reliability
Success rate across observed provider requests. This is historical sample data, not live health or an uptime guarantee.
Not enough observations
Last 7 days
There are not enough valid request observations to report a success rate.
Benchmarks
Dated third-party measurements with source links.
No verified benchmark snapshot is available for this model.
API Integration
Use the DOS API to integrate Gemini 3.5 Flash into your applications. Compatible with standard OpenAI SDKs for easy drop-in migration.
curl https://api.dos.ai/v1/chat/completions \
-H "Authorization: Bearer $DOS_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.5-flash",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
}'