Gemini 3.5 Flash

by Google

Latest Gemini Flash - frontier performance, standard tier

Parameters
Flash
Context Length
1M
Category
chat
Available Serverless

Run queries immediately, pay only for usage

$1.58in|$9.45out

Per 1M tokens

Try this modelView documentation

About this model

Latest Gemini Flash - frontier performance, standard tier

Capabilities

VisionLong context (1M)Tool callingStreaming

Providers

One model, several sources. DOS routes to the first that answers and falls back automatically, or you can pin a provider per request.

ProviderInput /1MOutput /1MContextServed from
GoogleAvailable
$1.58$9.451MUS

Model Details

Provider
Google
Model ID
gemini-3.5-flash
Parameters
Flash
Context Length
1M tokens
Category
chat

API Usage

Use the DOS API to integrate Gemini 3.5 Flash into your applications. Our API is compatible with OpenAI's client libraries for easy migration.

Model ID

gemini-3.5-flash

Python

python
from dos import DOS

client = DOS()

response = client.chat.completions.create(
    model="gemini-3.5-flash",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ]
)

print(response.choices[0].message.content)

cURL

bash
curl https://api.dos.ai/v1/chat/completions \
  -H "Authorization: Bearer $DOS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3.5-flash",
    "messages": [
      {"role": "user", "content": "Hello, how are you?"}
    ]
  }'

Node.js

javascript
import DOS from 'dos-ai';

const client = new DOS();

const response = await client.chat.completions.create({
  model: "gemini-3.5-flash",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ]
});

console.log(response.choices[0].message.content);