Gemini 2.0 Flash Lite

by Google

Google high-efficiency Flash model built for cost-sensitive, high-frequency multimodal workloads.

Parameters
Flash-Lite
Context Length
1M
Category
chat
Available Serverless

Run queries immediately, pay only for usage

$0.08in|$0.32out

Per 1M Tokens

Try this modelView documentation

About this model

Gemini 2.0 Flash Lite offers high-speed execution at the lowest cost in the Gemini 2.0 lineup, optimized for high-volume text and image processing.

Capabilities

Ultra-low costHigh throughputVision1M context

Use Cases

  • Batch processing
  • High-volume customer support
  • Classification
  • Summarization

Model Details

Provider
Google
Model ID
gemini-2.0-flash-lite
Parameters
Flash-Lite
Context Length
1M tokens
Category
chat

API Usage

Use the DOS API to integrate Gemini 2.0 Flash Lite into your applications. Our API is compatible with OpenAI's client libraries for easy migration.

Model ID

gemini-2.0-flash-lite

Python

python
from dos import DOS

client = DOS()

response = client.chat.completions.create(
    model="gemini-2.0-flash-lite",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ]
)

print(response.choices[0].message.content)

cURL

bash
curl https://api.dos.ai/v1/chat/completions \
  -H "Authorization: Bearer $DOS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-2.0-flash-lite",
    "messages": [
      {"role": "user", "content": "Hello, how are you?"}
    ]
  }'

Node.js

javascript
import DOS from 'dos-ai';

const client = new DOS();

const response = await client.chat.completions.create({
  model: "gemini-2.0-flash-lite",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ]
});

console.log(response.choices[0].message.content);