Gemini Embedding 001
ZDRAvailable ServerlessGoogle high-dimensional embedding model optimized for semantic search and Retrieval-Augmented Generation (RAG).
Providers
Different companies host the same model. DOS.AI routes your request to one of them based on the routing mode you pick — Balanced (price + speed), Nitro (fastest), or Exacto (highest tool-calling accuracy).
| Provider | Input /1M | Output /1M | Latency | Throughput | Uptime | Context | Retention |
|---|---|---|---|---|---|---|---|
Google | $0.02 | $0.00 | - | - | - | 2K | Not retainedZDR |
Pricing
Rates from the DOS.AI catalog for the billing units shown. Caching and provider discounts mean the price actually paid is often below the listed one.
Provider↕ | Effective in /M↕ | Effective out /M↕ | Cache read /M↕ | Latency↕ | Throughput↕ | Uptime↕ | Savings↕ |
|---|---|---|---|---|---|---|---|
Google | $0.02 | $0.00 | - | - | - | - | - |
Observed reliability
Success rate across observed provider requests. This is historical sample data, not live health or an uptime guarantee.
Not enough observations
Last 7 days
There are not enough valid request observations to report a success rate.
Benchmarks
Dated third-party measurements with source links.
No verified benchmark snapshot is available for this model.
API Integration
Use the DOS API to integrate Gemini Embedding 001 into your applications. Compatible with standard OpenAI SDKs for easy drop-in migration.
curl https://api.dos.ai/v1/embeddings \
-H "Authorization: Bearer $DOS_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-embedding-001",
"input": ["Text to embed"],
"encoding_format": "float"
}'