DoublewordDoubleword

Model Name

deepseek-ai/DeepSeek-V4-Pro

DeepSeek V4 Pro

  • Type: Generation
  • Capabilities: reasoning
  • Cache read: $0.13 per 1M input tokens (0.1× standard input price). See prompt caching.

Overview

DeepSeek V4-Pro is DeepSeek’s flagship open MoE model for advanced reasoning, coding, and agentic work. With 1.6T total parameters, 49B active parameters, and a 1M-token context window, it is designed for the hardest tasks in the V4 lineup: complex problem solving, knowledge-intensive workflows, and high-stakes coding or research use cases.

Reasoning efforts

  • Supported: none, minimal, low, medium, high, xhigh, max

See the reasoning effort guide for request examples.

Pricing

PriorityInput Tokens (per 1M)Cache Read Tokens (per 1M)Output Tokens (per 1M)
Realtime1$1.30$0.13$2.60
Async$0.98$0.10$1.95
Batch (24h)$0.65$0.07$1.30

Playground

Open this model in the Playground.

Footnotes

  1. Realtime availability is limited. Doubleword is primarily a batch API.