← All model brands · BD LLM Token Hub
LLM knowledge · BD Token Hub

DeepSeek — frontier reasoning at commodity prices

DeepSeek’s V4 line pairs flagship reasoning (V4 Pro) with a budget agent workhorse (V4.1 Flash) — open weights, 1M context, and off-peak pricing that undercuts most of the market.

V4 Pro left preview 12–13 Aug 2026 and runs OpenAI Codex natively.

Who they are

DeepSeek is the Chinese lab that reset global price expectations with its open-weight MoE models. V4 Pro carries 1.6T total / 49B active parameters with a 1M context window; V4.1 Flash runs 284B / 13B for high-volume agent work.

Model lineup

ModelWhat it is forVendor list price
DeepSeek V4 Proflagship · GA 13 Aug 2026 · 1M ctxReasoning and agentic MoE (1.6T/49B active) — 80.6% SWE-bench Verified, runs Codex nativelyUS$0.66 / US$1.98 per 1M off-peak (≈ RM2.69 / RM8.08) · peak US$1.32 / US$3.96
DeepSeek V4.1 Flashbudget agent · GA 10 Sep 2026 · 1M ctx284B/13B active for high-volume agent tasks, with native visionUS$0.15 / US$0.60 per 1M off-peak (≈ RM0.61 / RM2.45) · peak US$0.30 / US$1.20

Last updated — Off-peak/peak billing since 17 Aug 2026 (peak = 2x off-peak, on input and output alike). Approx RM at 1 USD ≈ RM4.08 (spot, 19 Sep 2026). Cache-hit input US$0.003–0.044 /1M off-peak. On BD LLM Token Hub today. Figures are vendor public list prices in USD per 1M tokens, standard tier — no batch, cache, promo or peak-hour discount applied. Vendors change list prices without notice: confirm on the vendor’s own pricing page before budgeting.

Historical models — for reference

Discontinued or fully superseded models, kept for historical reference. Active models — even previous-generation ones still on sale — always stay in the lineup above.

ModelEra & what it was forHistoric list price
DeepSeek R1Jan 2025 · reasoning eraThe open reasoning model that shook the marketUS$0.55 / US$2.19 per 1M (≈ RM2.24 / RM8.94) — historic list
DeepSeek V3Dec 2024The original V3 general MoEUS$0.27 / US$1.10 per 1M (≈ RM1.10 / RM4.49) — historic list

Latest news

Recent announcements with outlet links (EN first, CN where available).

13 Aug 2026

V4 Pro exits preview — runs OpenAI Codex natively

Flagship reasoning model goes GA at ~8x cheaper than comparable closed models, with native Codex support.

EN · Unite.AI ↗EN · AlphaSignal ↗
17 Aug 2026

Peak / off-peak billing introduced

DeepSeek moved to time-based pricing — off-peak is half of peak on both input and output, so agent schedules can be shaped around cost. Current per-model rates are in the table above.

EN · DeepSeek pricing ↗EN · AI Pricing Guru ↗
10 Sep 2026

Flash tier becomes V4.1 Flash — at lower prices

The budget agent model is now DeepSeek-V4.1-Flash (model id deepseek-flash) with native vision, billed off-peak at US$0.15/M input and US$0.60/M output. The retired deepseek-v4-flash name routes to it at the Flash price.

EN · DeepSeek API docs ↗EN · AI Pricing Guru ↗

Want DeepSeek models behind one API key?

Ask about DeepSeek on Token Hub One key gets you V4 Pro and V4.1 Flash on BD LLM Token Hub — ask us for today’s rates.
More from Big Domain

You might also need…

One partner for every digital layer – explore what else we build, host and grow for Malaysian businesses.