Models (24)
Featured Models
DOS.AI Auto
Smart routing - automatically picks the best model for your request. Free for simple tasks, paid models for complex ones.
DOS.AI
Ultra-efficient MoE model — 35B total, 3B active parameters. Fast inference at near-8B cost with 70B-class quality.
DeepSeek V4 Flash
1M-context fast tier replacing DeepSeek V3
Gemini 3.1 Pro
Google's most advanced reasoning model for complex tasks
GPT-5.5
OpenAI flagship — replaces GPT-5.4 at top tier with native reasoning
Claude Opus 4.8
All Models
DOS.AI Auto
Smart routing - automatically picks the best model for your request. Free for simple tasks, paid models for complex ones.
DOS.AI
Ultra-efficient MoE model — 35B total, 3B active parameters. Fast inference at near-8B cost with 70B-class quality.
DeepSeek V4 Flash
1M-context fast tier replacing DeepSeek V3
GPT-5.4 Nano
Cheapest GPT-5.4-class model for simple high-volume tasks
Gemini 3.1 Flash-Lite
Fastest and most cost-efficient Gemini 3 model
Gemini 3.5 Flash-Lite
Fast, cost-effective Flash-Lite (GA). 1M context. Closest Gemini peer to dos-ai (Qwen3.6-35B-A3B). [PROMO] Free when used BY a DOSClaw agent on a paid plan (agent traffic only - direct API usage is billed normally). Limited-time.
Gemini 2.5 Flash
Previous-generation Flash, best price-performance. GA and stable.
Qwen 3.7 Plus
Qwen 3.7 Plus - multimodal (text+image), 1M context
Gemini 3.1 Flash Live
Real-time voice and dialogue model
GPT-5.4 Mini
Strong mini model for coding, computer use, and sub-agents
Claude Haiku 4.5
Fastest and most compact Claude model
Grok 4.3
xAI flagship reasoning model - 1M context, text+image (replaces grok-4.1-fast)
Gemini 3.5 Flash
Latest Gemini Flash - frontier performance, standard tier
Gemini 3.6 Flash
Google latest balanced Flash model (launched 2026-07-21). 1M context, 64k output.
DeepSeek V4 Pro
Near-frontier quality at ~1/6 the cost of Opus 4.7 / GPT-5.5
Grok 4.20
xAI flagship reasoning model with 2M context
Gemini 3.1 Pro
Google's most advanced reasoning model for complex tasks
Qwen 3.7 Max
Qwen 3.7 Max tier - 1M context (standard price; OpenRouter shows a 50%-off promo)
Claude Sonnet 4.6
Fast, intelligent model for everyday tasks
GPT-5.5
OpenAI flagship — replaces GPT-5.4 at top tier with native reasoning
Claude Opus 4.8
Wan 2.7 Text-to-Video
Video generation from text prompt via Alibaba Wan 2.7. Duration 2-15s, 1080P, native audio. Pricing: per 1000 seconds.
Wan 2.7 Image-to-Video
Video generation from image + text prompt via Alibaba Wan 2.7. Pricing: per 1000 seconds.
Qwen3 Embedding 4B
Self-hosted multilingual text embedding (2560 dims), strong on Vietnamese. Priced at our OpenRouter fallback cost.
Ready to get started?
Start building with $10 in free credits. No credit card required.