Gemini 2.0 Flash Thinking

by Google

Google experimental reasoning model providing explicit chain-of-thought steps alongside fast multimodal responses.

Parameters
Flash Thinking
Context Length
1M
Category
chat
Available Serverless

Run queries immediately, pay only for usage

$0.11in|$0.42out

Per 1M Tokens

Try this modelView documentation

About this model

Gemini 2.0 Flash Thinking is Google's dedicated reasoning model that displays explicit thinking steps before answering, allowing users and developers to inspect the reasoning chain.

Capabilities

Visible chain-of-thoughtDeep math & reasoningVision1M context

Use Cases

  • Complex problem breakdown
  • Math & logic verification
  • Code debugging with rationale

Model Details

Provider
Google
Model ID
gemini-2.0-flash-thinking
Parameters
Flash Thinking
Context Length
1M tokens
Category
chat

API Usage

Use the DOS API to integrate Gemini 2.0 Flash Thinking into your applications. Our API is compatible with OpenAI's client libraries for easy migration.

Model ID

gemini-2.0-flash-thinking

Python

python
from dos import DOS

client = DOS()

response = client.chat.completions.create(
    model="gemini-2.0-flash-thinking",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ]
)

print(response.choices[0].message.content)

cURL

bash
curl https://api.dos.ai/v1/chat/completions \
  -H "Authorization: Bearer $DOS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-2.0-flash-thinking",
    "messages": [
      {"role": "user", "content": "Hello, how are you?"}
    ]
  }'

Node.js

javascript
import DOS from 'dos-ai';

const client = new DOS();

const response = await client.chat.completions.create({
  model: "gemini-2.0-flash-thinking",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ]
});

console.log(response.choices[0].message.content);