DeepSeek

DeepSeek supplies hosted language models with token-based billing. These estimates use peak-hour, uncached rates; off-peak and cache-hit requests can cost less.

Categories
LLM
Models and endpoints
2
Guide reviewed
Oct 8, 2026

Models, translated into money

Choose a budget and compare what each model delivers. Every estimate includes its assumptions and pricing sources.

TaskBudget · USD
Compare2 models
Task
Budget USD

DeepSeek-V4.1-Flash (peak, uncached)

$0.30 input · $1.20 output / 1M tokens

What $10 gets you

About 8,333chats

2,000 input + 500 output tokens per chat

Where it fits

Based on independent evidence
ChatNo data

Not enough current evidence to assess this task.

SupportNo data

Not enough current evidence to assess this task.

CodingNo data

Not enough current evidence to assess this task.

DocumentsNo data

Not enough current evidence to assess this task.

ResearchNo data

Not enough current evidence to assess this task.

ByeTokens

DeepSeek-V4.1-Flash (peak, uncached)


  • Your budget$10
  • Buys about8,333 chats
  • WorkloadTypical chat2,000 input + 500 output tokens per request

  • Price examples
    100 chats$0.1200
  • 1,000 chats$1.20
  • 10,000 chats$12.00

Catalog 2026-10-08-7b05c4ce
Each line is a separate example.

Calculate with DeepSeek-V4.1-Flash (peak, uncached)Use in calculator
How we calculated thisVerified Oct 8
Typical chat assumption2,000 input · 500 output tokens per request

Official pricing

deepseek-flash-peak-uncached · Current · Standard billing

Standard text
$0.30 input · $1.20 output / 1M tokens

Verified Oct 8, 2026

Model details

Model ID
deepseek-flash
Tool calling
true
Structured output
unknown
context Window
1,000,000
max Output Tokens
384,000
modality
text
model Version
DeepSeek-V4.1-Flash
processing Tier
standard
cached Input
No
pricing Period
peak

Task evidence

General chat: No current score

Support chat: No current score

Coding help: No current score

Document work: No current score

Research agent: No current score

Catalog 2026-10-08-7b05c4ce

DeepSeek-V4-Pro-0813 (peak, uncached)

$1.32 input · $3.96 output / 1M tokens

What $10 gets you

About 2,164chats

2,000 input + 500 output tokens per chat

Where it fits

Based on independent evidence
ChatNo data

Not enough current evidence to assess this task.

SupportNo data

Not enough current evidence to assess this task.

CodingNo data

Not enough current evidence to assess this task.

DocumentsNo data

Not enough current evidence to assess this task.

ResearchNo data

Not enough current evidence to assess this task.

ByeTokens

DeepSeek-V4-Pro-0813 (peak, uncached)


  • Your budget$10
  • Buys about2,164 chats
  • WorkloadTypical chat2,000 input + 500 output tokens per request

  • Price examples
    100 chats$0.4620
  • 1,000 chats$4.62
  • 10,000 chats$46.20

Catalog 2026-10-08-7b05c4ce
Each line is a separate example.

Calculate with DeepSeek-V4-Pro-0813 (peak, uncached)Use in calculator
How we calculated thisVerified Oct 8
Typical chat assumption2,000 input · 500 output tokens per request

Official pricing

deepseek-v4-pro-peak-uncached · Current · Standard billing

Standard text
$1.32 input · $3.96 output / 1M tokens

Verified Oct 8, 2026

Model details

Model ID
deepseek-v4-pro
Tool calling
true
Structured output
unknown
context Window
1,000,000
max Output Tokens
384,000
modality
text
model Version
DeepSeek-V4-Pro-0813
processing Tier
standard
cached Input
No
pricing Period
peak

Task evidence

General chat: No current score

Support chat: No current score

Coding help: No current score

Document work: No current score

Research agent: No current score

Catalog 2026-10-08-7b05c4ce

About DeepSeek

Coverage, billing details, and what to check before you build.

DeepSeek is included as a direct hosted API provider. The two catalog entries use the current endpoint names, deepseek-flash and deepseek-v4-pro, rather than treating retired model aliases as additional products. Their model versions and official sources appear in the offer details above.

What the estimate covers

The catalog models text input and output tokens for language calls. DeepSeek documents tool calling, JSON output, thinking and non-thinking modes for both endpoints. Both have a one-million-token context window and a maximum output of 384,000 tokens. Flash also supports vision, but this calculator's text workload does not estimate image-input costs.

Search and page retrieval remain separate services in a stack. The calculator does not turn tool-calling support into an assumption that external search, scraping, or other tool usage is included in the language-model price.

Peak rates, without assumed cache savings

Each offer uses the published peak-hour rate for input tokens that miss the cache, plus the peak output rate. The model names explicitly say “peak, uncached” so this assumption stays visible when comparing providers or opening a locked scenario.

DeepSeek also publishes lower off-peak and cache-hit prices. Those discounts are excluded here because the current form does not collect request timing or cache-hit proportions. This is a conservative token-cost baseline, not a prediction that every request will be billed at the peak rate. Check the official pricing page for the schedule and discounted paths before reconciling an invoice.

Match the workload to your application

Select an offer to open the calculator with that model locked. Describe monthly activity, or use exact mode when you know the number of calls and tokens per call. Input and output tokens are charged separately, and the engine checks the combined context and output limits before returning a complete estimate.

The preset chat, support, coding, and document quantities are token assumptions. They do not establish task quality or guarantee that an application will use the same number of tokens. Adjust the workload to match measured requests, including the billable output reported by the API.

Capability and evidence boundaries

JSON output does not by itself establish support for a strict JSON schema. The catalog leaves that capability unknown rather than promising schema enforcement. Requiring structured output in the calculator therefore excludes these entries until that requirement is verified.

No independent task-quality score is added as part of this pricing entry. A known price can support a cost comparison while evidence about suitability remains incomplete. Use your own representative requests to evaluate behavior, and inspect the displayed verification date and sources when planning a budget. Taxes, infrastructure, retries, and development costs are outside the estimate.