Home → Pricing → deepseek-v4-pro

DeepSeek V4 Pro API pricing

deepseek-v4-pro

DeepSeek's flagship generation. Heavily used here for long-context work — it accounts for more tokens than any other model on the platform. Dated snapshots such as deepseek-v4-pro-0813 are the same model and are billed the same way.

DeepSeek You pay 20% of list price OpenAI-compatible Pay as you go
List price · input
$1.32
per 1M tokens · output $3.96
You pay · input
$0.264
per 1M tokens · output $0.792
You save 80% on every call. Same model, same context window, same speed — the only thing that changes is your bill.

Price breakdown

Per 1M tokensList priceAPICLANYou save
Input$1.32$0.26480%
Output$3.96$0.79280%

Cached input is billed at a fraction of the input rate, so agent tools that resend the same context — Codex, Claude Code, Cline — cost considerably less in practice than the raw token count suggests.

Cached input is billed at a fraction of the input rate, which matters here more than most: the traffic we see on this model is dominated by very large prompts.

Work out your own cost

Put in the token counts from a typical request. Both columns update as you type.

At list price
—
At APICLAN
—

What a real session costs

A coding session through Codex on this platform used 90,459 tokens across five requests — roughly seventeen minutes of back-and-forth over a codebase. At list price that is $0.178. The customer paid $0.036.

That is real billing data, not an estimate. Heavy context reuse is exactly where the saving compounds: the same work, one fifth of the invoice.

How to call deepseek-v4-pro

The API is OpenAI-compatible. Point your base URL at APICLAN and pass the model name — nothing else in your code changes.

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_KEY",
    base_url="https://apiclan.us/v1",
)

resp = client.chat.completions.create(
    model="deepseek-v4-pro",
    messages=[{"role": "user", "content": "Hello"}],
)

Using Claude Code or the Anthropic SDK? Those clients append /v1/messages themselves, so their base URL is https://apiclan.us with no /v1. Full setup for every client is in the Quickstart.

Other DeepSeek models

ModelInput / 1MOutput / 1M
deepseek-v4-flash$0.088$0.264

Prices above are what you pay, not list price. Cheaper models in the same family often handle routine work at a fraction of the cost — worth testing before defaulting to the largest one.

Start using it

No subscription, no monthly minimum, no sales call. Top up with USDT and spend what you use — 1 USDT gives you 2 credits of API balance.

Read the 30-second quickstart

Pricing verified 2026-10-01 · from the published list price. List prices change; this page is regenerated automatically from live billing data.

Language: English · 简体中文 · 繁體中文 · 日本語 · 한국어 · Deutsch · Français · Español · Türkçe · Português · Italiano · Bahasa Indonesia · Tiếng Việt · Bahasa Melayu · Polski · Čeština · فارسی · Русский