DOSDOS
Pricing
Get started
Models/DOS.AI

DOS.AI

New

by DOS.AI

Ultra-efficient MoE model — 35B total, 3B active parameters. Fast inference at near-8B cost with 70B-class quality.

Parameters
35B MoE (3B active)
Context Length
128K
Category
chat
Available Serverless

Run queries immediately, pay only for usage

$0.07in|$0.50out

Per 1M Tokens

Try this modelView documentation

About this model

DOS.AI is a Mixture-of-Experts model (35B total, 3B active per token) served on our own GPUs in Vietnam. It delivers 70B-class quality at 8B-class speed and cost, with strong multilingual, tool-calling, and coding performance. At $0.07 / $0.50 per million tokens it is roughly half the price of comparable providers, with full data residency in Vietnam.

Capabilities

MultilingualTool callingCode generationStructured outputStreaming

Use Cases

  • AI agents
  • Customer support
  • Content generation
  • Code assistance

Model Details

Provider
DOS.AI
Model ID
dos-ai
Parameters
35B MoE (3B active)
Context Length
128K tokens
Category
chat

API Usage

Use the DOS API to integrate DOS.AIinto your applications. Our API is compatible with OpenAI's client libraries for easy migration.

Model ID

dos-ai

Python

python
from dos import DOS

client = DOS()

response = client.chat.completions.create(
    model="dos-ai",
    messages=[
        {"role": "user", "content": "Hello, how are you?"}
    ]
)

print(response.choices[0].message.content)

cURL

bash
curl https://api.dos.ai/v1/chat/completions \
  -H "Authorization: Bearer $DOS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "dos-ai",
    "messages": [
      {"role": "user", "content": "Hello, how are you?"}
    ]
  }'

Node.js

javascript
import DOS from 'dos-ai';

const client = new DOS();

const response = await client.chat.completions.create({
  model: "dos-ai",
  messages: [
    { role: "user", content: "Hello, how are you?" }
  ]
});

console.log(response.choices[0].message.content);
View full API reference

Related Models

DOS.AI

DOS.AI Auto

Smart routing - automatically picks the best model for your request. Free for simple tasks, paid models for complex ones.

Usage-based
DeepSeek

DeepSeek V4 Flash

1M-context fast tier replacing DeepSeek V3

$0.15 / 1M tokens
OpenAI

GPT-5.4 Nano

Cheapest GPT-5.4-class model for simple high-volume tasks

$0.21 / 1M tokens
DOSDOS

AI infrastructure for everyone. Inference, agents, and safety - all in one platform.

Product

  • Models
  • Pricing
  • API Inference
  • DOSClaw

Developers

  • Documentation
  • API Reference
  • Status

DOS Ecosystem

  • DOSafe
  • DOS.Me
  • DOScan
  • DOSwap
  • MetaDOS

Company

  • About
  • Contact
  • Careers
  • Privacy
  • Terms

© 2026 All rights reserved.