GPT-OSS 120B
by deepinfra
OpenAI open-weight 120B, strong general reasoning at a fraction of hosted GPT pricing.
Run queries immediately, pay only for usage
Per 1M tokens
About this model
OpenAI open-weight 120B, strong general reasoning at a fraction of hosted GPT pricing.
Capabilities
Providers
One model, several sources. DOS routes to the first that answers and falls back automatically, or you can pin a provider per request.
| Provider | Input /1M | Output /1M | Context | Served from |
|---|---|---|---|---|
DeepInfraAvailablebfloat16 | $0.05 (cheapest) | $0.18 (cheapest) | 128K | US |
Together AIAvailable | $0.16 | $0.63 | 128K | US |
Model Details
- Provider
- Model ID
- gpt-oss-120b
- Parameters
- Context Length
- 128K tokens
- Category
- chat
API Usage
Use the DOS API to integrate GPT-OSS 120B into your applications. Our API is compatible with OpenAI's client libraries for easy migration.
Model ID
gpt-oss-120bPython
from dos import DOS
client = DOS()
response = client.chat.completions.create(
model="gpt-oss-120b",
messages=[
{"role": "user", "content": "Hello, how are you?"}
]
)
print(response.choices[0].message.content)cURL
curl https://api.dos.ai/v1/chat/completions \
-H "Authorization: Bearer $DOS_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-oss-120b",
"messages": [
{"role": "user", "content": "Hello, how are you?"}
]
}'Node.js
import DOS from 'dos-ai';
const client = new DOS();
const response = await client.chat.completions.create({
model: "gpt-oss-120b",
messages: [
{ role: "user", content: "Hello, how are you?" }
]
});
console.log(response.choices[0].message.content);Related Models
Auto
DOS.AI Smart Router classifies each request by complexity. Simple and Medium requests use DOS.AI; Complex requests can use an eligible paid route configured in the live catalog.
Qwen3.8-27B
DOS.AI hosted Qwen3.8-27B for reasoning, coding, tool use, and multilingual text generation.
DOS
DOS curated meta-model for DOSClaw and first-party DOS products.