Back to models

DeepSeek V4 Pro

BytePlus•

Near-frontier quality at ~1/6 the cost of Opus 4.7 / GPT-5.5

Capabilities:Advanced reasoningMath and logicCode generationLong context (1M)
Modalities
Input / Output Price
$1.32/$3.96/ 1M
Context Length
1M tokens
Released
Apr 24, 2026

Providers

Different companies host the same model. DOS.AI routes your request to one of them based on the routing mode you pick — Balanced (price + speed), Nitro (fastest), or Exacto (highest tool-calling accuracy).

Reasoning effortⓘAll
ProviderInput /1MOutput /1MCache read /M
Latency
UptimeContext
Cloudflare
$1.32 (cheapest)$3.96$0.044--1M
BytePlus
$1.40$2.80 (cheapest)$0.14--1M
Alibaba Cloud
$2.40$4.80$0.20--1M

Pricing

Rates from the DOS.AI catalog for the billing units shown. Caching and provider discounts mean the price actually paid is often below the listed one.

Weighted Averagei
Weighted Avg Input Price
$1.707/ 1M tokens
Equal to list price
Weighted Avg Output Price
$3.853/ 1M tokens
Equal to list price
Cache read
$0.044/ 1M tokens
Rate for cached input tokens
Price History3 verified endpoints
Weighted Average($1.71)
$0.00$0.42$0.85$1.2830d ago24d ago18d ago12d ago6d agoToday
Provider↕
Effective in /M↕
Effective out /M↕
Cache read /M↕
Latency↕
Uptime↕
Savings↕
BytePlus
$1.40$0.64
$2.80$0.14---54%
Cloudflare
$1.32$0.55
$3.96$0.04---58%
Alibaba Cloud
$2.40$1.08
$4.80$0.20---55%
Based on settled billing for the last 30d.Effective pricing represents the average price paid for requests to this endpoint, factoring in caching, discounts, and tiered pricing. Listed pricing shows the posted provider list prices.

Observed reliability

Success rate across observed provider requests. This is historical sample data, not live health or an uptime guarantee.

Not enough observations

Last 7 days

There are not enough valid request observations to report a success rate.

Benchmarks

Dated third-party measurements with source links.

No verified benchmark snapshot is available for this model.

API Integration

Use the DOS API to integrate DeepSeek V4 Pro into your applications. Compatible with standard OpenAI SDKs for easy drop-in migration.

</>Run inference
curl https://api.dos.ai/v1/chat/completions \
  -H "Authorization: Bearer $DOS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-pro",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }'