PeakPrice
Live cost tool

DeepSeek Peak Pricing Calculator

Calculate DeepSeek V4 Flash and V4 Pro costs at peak or off-peak rates, see the schedule in your timezone, and estimate what a month of API traffic could cost.

Rates from DeepSeek's official API documentation. No sign-up, API key, or prompt data required.

Checking live pricing…

Using your local timezone

Off-peak begins in

--:--:--

Cost inputs

Estimate your API bill

Model
Request time

Peak windows are fixed in UTC. We apply daylight-saving changes automatically.

tokens
tokens
50%

DeepSeek reports cache-hit and cache-miss input tokens separately. Use your observed average when available.

requests

Rates verified Aug 14, 2026 · New schedule effective Aug 16 at 16:00 UTC

Peak windows: 01:00–04:00 and 06:00–10:00 UTC

7 hours

of peak pricing each day

17 hours

available at off-peak rates

50% lower

off-peak rates than peak rates

Pricing guide

The numbers behind your estimate

A practical guide to DeepSeek's time-based API rates, context caching, monthly budgets, and workloads that can safely move off-peak.

01

How the DeepSeek Peak Pricing Calculator Works

The DeepSeek Peak Pricing Calculator turns token estimates into a practical cost. Choose V4 Flash or V4 Pro, enter input and output tokens, and set an estimated cache hit ratio. For a single request, it applies the rate in effect at your selected time and shows the same workload at both peak and off-peak prices.

Monthly mode uses average tokens, requests per day, billing days, and the share of traffic expected during peak hours. It calculates the two contributions separately, then compares the result with an all-off-peak schedule. Estimates use US dollars per 1 million tokens. They are planning figures, not a replacement for API usage records or a DeepSeek bill.

02

DeepSeek Peak and Off-Peak Hours

DeepSeek lists two daily peak windows: 01:00–04:00 UTC and 06:00–10:00 UTC. The other 17 hours are off-peak. At the boundaries, 01:00 and 06:00 are peak; 04:00 and 10:00 begin off-peak pricing. The underlying schedule always stays in UTC.

Peak and off-peak billing takes effect at 16:00 UTC on August 16, 2026. The live status classifies time in UTC, then presents the current time and next switch in your selected IANA timezone. This accounts for daylight saving time where it applies. If detection fails, the calculator uses UTC and tells you.

03

DeepSeek V4 Flash and V4 Pro Pricing

The official table covers deepseek-v4-flash (DeepSeek-V4-Flash-0731) and deepseek-v4-pro (DeepSeek-V4-Pro-0813). Each has separate cache-hit input, cache-miss input, and output rates. Under the new schedule, every off-peak rate is half of its matching peak rate. V4 Flash costs less across all three categories, but quality, latency, and workload needs still belong in a production decision.

These figures apply to calls billed directly by DeepSeek. A cloud marketplace, gateway, reseller, or aggregator may set a different markup, currency, discount, or schedule. Use that provider's documentation for third-party traffic.

DeepSeek V4 peak and off-peak API rates

USD per 1 million tokens · Effective from 16:00 UTC on August 16, 2026

ModelPricing periodInput: cache hitInput: cache missOutput
deepseek-v4-flashOff-Peak$0.007$0.22$0.66
Peak$0.014$0.44$1.32
deepseek-v4-proOff-Peak$0.022$0.66$1.98
Peak$0.044$1.32$3.96
04

Cache Hit vs Cache Miss Input Tokens

DeepSeek's context caching is enabled by default. A later request can receive a cache hit when it fully reuses a persisted prefix; other input is billed as a cache miss. Output tokens have their own price and are not affected by the input cache ratio.

The calculator splits input using the ratio you enter. A 60% ratio treats 60% as cache hits and 40% as misses. This is an estimate, not an API switch. Stable prefixes and multi-turn conversations may improve reuse, while frequently changing content may not. For a tighter forecast, use representative cache hit and miss counts from real API traffic.

05

How to Estimate Your Monthly DeepSeek API Bill

Start with a normal request: average input, average output, and the cache behavior you see in practice. Apply daily volume and billing days, then choose a peak traffic share. If calls are evenly spread across the day, 7 of 24 hours are peak, so 29.2% is a useful default—not a universal assumption.

Interactive products may follow their audience's working day, while global systems can be flatter. Measure your distribution with broad, non-identifying time buckets. Test expected volume, a growth case, and a cache-miss-heavy case to understand the range your budget may need to absorb.

06

How Much Can Off-Peak Scheduling Save?

Moving work from peak to off-peak cuts the published token rate by 50%. The savings card compares your selected mix with the same volume entirely off-peak. It does not count engineering time, queueing infrastructure, or the cost of delay.

Treat the result as an upper bound for the traffic represented by your inputs. Customer-facing and time-sensitive work may need to run immediately. Better candidates have a completion deadline but no need to start at once.

07

Common Workloads You Can Move Off-Peak

Batch evaluation, test generation, offline summarization, search indexing, report drafting, classification, and nightly code analysis are often queue-friendly. A scheduler can hold jobs for an off-peak window and retry safely. Keep user-visible work on the path that meets its latency target, then optimize the flexible remainder.

Before moving a job, record its deadline, token profile, and failure policy. Schedule against UTC because DeepSeek defines the windows there. Recalculate after a model change, prompt redesign, or shift in cache performance. Reducing tokens on every request can matter as much as rescheduling a limited batch.

FAQ

DeepSeek pricing questions, answered

Quick answers about peak windows, weekends, caching, model coverage, and what a monthly estimate can—and cannot—predict.

01Is DeepSeek peak pricing active now?

DeepSeek's peak and off-peak billing takes effect at 16:00 UTC on August 16, 2026. Once active, the live status checks UTC and marks 01:00 to 04:00 and 06:00 to 10:00 UTC as peak hours.

02What are DeepSeek peak hours in my timezone?

Select your IANA timezone to convert the UTC windows for that date. The windows stay fixed in UTC, while local display times can shift with daylight saving time.

03Are weekends priced differently?

DeepSeek identifies daily peak windows and says all other hours are off-peak. It lists no separate weekend schedule, so the calculator applies the same UTC windows every day.

04Does the calculator support both DeepSeek V4 models?

Yes. It covers deepseek-v4-flash and deepseek-v4-pro from DeepSeek's current API table. It does not estimate discontinued names, self-hosted deployments, or third-party rates.

05How does a cache hit change the estimate?

Cache-hit input uses a lower published rate than cache-miss input. Your chosen ratio applies only to input; output remains separate. Use observed API usage when possible because actual caching depends on reusable prefixes.

06Can the monthly estimate predict my exact invoice?

No. It models the usage and traffic mix you provide. Actual tokens, caching, timing, retries, pricing changes, and billing rules can change the total. DeepSeek's invoice is the source of truth.

07Is off-peak always the best time to call the API?

Off-peak has the lower rate, but timing is only one constraint. Run urgent work when needed; use the savings estimate to evaluate delay-tolerant jobs.

A better cost baseline

Plan the cost, then verify the bill

Use the DeepSeek Peak Pricing Calculator whenever traffic volume, prompts, caching, or scheduling changes. It gives teams a shared cost model before code ships and makes the peak-versus-off-peak choice concrete. Check DeepSeek's official pricing page regularly, compare estimates with actual usage, and treat the provider's invoice as final.

Read the methodology, privacy notes, and limitations