Gemini 1.5 Flash

by Google

Google lightweight, fast, and cost-effective multimodal model for high-frequency tasks with 1M context.

Parameters
Flash (1M)
Context Length
1M
Category
chat
Available Serverless

Run queries immediately, pay only for usage

$0.08in|$0.32out

Per 1M Tokens

Try this modelView documentation

About this model

Google lightweight, fast, and cost-effective multimodal model for high-frequency tasks with 1M context.

Capabilities

Fast inference1M contextMultimodalStreaming

Model Details

Provider
Google
Model ID
gemini-1.5-flash
Parameters
Flash (1M)
Context Length
1M tokens
Category
chat

API Usage

Use the DOS API to integrate Gemini 1.5 Flash into your applications. Our API is compatible with OpenAI's client libraries for easy migration.

Model ID

gemini-1.5-flash

Python

python
from dos import DOS

client = DOS()

response = client.chat.completions.create(
    model="gemini-1.5-flash",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ]
)

print(response.choices[0].message.content)

cURL

bash
curl https://api.dos.ai/v1/chat/completions \
  -H "Authorization: Bearer $DOS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-1.5-flash",
    "messages": [
      {"role": "user", "content": "Hello, how are you?"}
    ]
  }'

Node.js

javascript
import DOS from 'dos-ai';

const client = new DOS();

const response = await client.chat.completions.create({
  model: "gemini-1.5-flash",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ]
});

console.log(response.choices[0].message.content);