AI Cost Intelligence, Verified Daily.
Skip to content
Pricing updates6 min readUpdated Sep 6, 2026

DeepSeek V3.2 Exp Price Change | AIOPLY Cost Report

Our pricing desk analyzes the latest DeepSeek V3.2 Exp price change recorded on 2026-08-29, calculating real monthly API cost savings and evaluating cheaper alternatives across 36 tracked models.

DeepSeek V3.2 Exp Price cost chart in the AIOPLY house style
DeepSeek V3.2 Exp Price cost chart in the AIOPLY house style

Key takeaways

  • On 2026-08-29, DeepSeek reduced the V3.2 Exp output token price by 4.8 percent from $0.42 to $0.40 per 1M tokens.
  • Input pricing for DeepSeek V3.2 Exp fell by 3.9 percent from $0.28 to $0.269 per 1M tokens.
  • A monthly API workload of 10M input and 2M output tokens sees costs fall from $3.64 to $3.49.
  • DeepSeek V3.2 Exp output pricing remains significantly lower than the $3.90 per 1M token median across 36 tracked models.
  • Teams needing extreme low cost or larger context windows should consider DeepSeek V4 Flash 0731, which offers 1311k tokens at $0.045 input and $0.09 output per 1M tokens.

What changed in the DeepSeek V3.2 Exp price change?

On 2026-08-29, automated pricing logs recorded a rate reduction across both input and output tokens for DeepSeek V3.2 Exp. The deepseek v3.2 exp price change lowers output token charges from $0.42 to $0.40 per 1M tokens, representing a 4.8 percent reduction while input token rates moved from $0.28 to $0.269 per 1M tokens, a 3.9 percent decrease.

This update was applied directly by provider DeepSeek across its API infrastructure, leaving the model's 164k token context window unchanged. While single digit rate cuts do not completely transform software unit economics, they alter developer calculations when comparing models for high output pipelines.

Our automated pricing monitor captured this shift during routine database indexing, allowing us to evaluate how the new rates impact real engineering budgets compared to other models currently tracked in our system. The adjustments reflect ongoing price adjustments among foundation model providers attempting to maintain usage volume across experimental model tiers.

What are the exact input and output rates for DeepSeek V3.2 Exp?

The revised rate card for DeepSeek V3.2 Exp sets input pricing at $0.269 per 1M tokens and output pricing at $0.40 per 1M tokens. This price structure represents a drop from the previous rate card of $0.28 per 1M input tokens and $0.42 per 1M output tokens recorded prior to 2026-08-29.

When analyzing deepseek v3.2 exp pricing, the ratio between input and output costs remains relatively narrow compared to industry norms. Output tokens cost roughly 1.48 times as much as input tokens ($0.40 versus $0.269). Prior to the change, output tokens cost 1.50 times input tokens ($0.42 versus $0.28).

This narrow ratio benefits workloads with high output volume relative to prompt size, such as long text generation or detailed step-by-step reasoning outputs. Understanding this pricing structure allows developers to project expenses accurately when processing large batches of generation requests within the 164k token context boundary.

ModelProviderInput Rate (per 1M)Output Rate (per 1M)Context Window
DeepSeek V3.2 Exp (New)DeepSeek$0.269$0.400164k tokens
DeepSeek V3.2 Exp (Previous)DeepSeek$0.280$0.420164k tokens
DeepSeek V4 Flash 0731DeepSeek$0.045$0.0901311k tokens
Qwen3.7 FlashQwen$0.030$0.130Not specified
Nemotron 3.5 LightningNVIDIA$0.080$0.200Not specified
Token Pricing Comparison Across Selected Current Models

How does DeepSeek V3.2 Exp compare to Flash alternatives?

Lighter Flash models available in our database offer substantially lower token costs than DeepSeek V3.2 Exp. For example, DeepSeek V4 Flash 0731 charges $0.045 per 1M input tokens and $0.09 per 1M output tokens while offering a 1311k token context window.

Engineering teams seeking maximum api cost savings can evaluate three primary alternatives recorded in our current tracking dataset:

DeepSeek V4 Flash 0731 from provider DeepSeek charges $0.045 per 1M input tokens and $0.09 per 1M output tokens. In addition, DeepSeek V4 Flash 0731 features a context window of 1311k tokens, which is the widest context in the current dataset. Compared to DeepSeek V3.2 Exp, V4 Flash 0731 lowers input expenses by 83.3 percent and output expenses by 77.5 percent while offering over seven times the context capacity.

Qwen3.7 Flash from provider Qwen charges $0.03 per 1M input tokens and $0.13 per 1M output tokens. Nemotron 3.5 Lightning from provider NVIDIA charges $0.08 per 1M input tokens and $0.20 per 1M output tokens. Both options deliver significantly lower token rates across input and output billing categories than the revised DeepSeek V3.2 Exp rate card, providing alternatives for budget constrained pipelines.

  • DeepSeek V4 Flash 0731: $0.045 input / $0.09 output per 1M tokens (1311k context)
  • Qwen3.7 Flash: $0.03 input / $0.13 output per 1M tokens
  • Nemotron 3.5 Lightning: $0.08 input / $0.20 output per 1M tokens

What is the exact monthly bill impact across volume tiers?

A workload sending 10M input and 2M output tokens a month now costs $3.49, against $3.64 before the change. Following this deepseek v3.2 exp price change, an organization running this volume profile saves exactly $0.15 per month, equivalent to a 4.1 percent overall bill reduction.

The underlying arithmetic breaks down clearly across token types:

Scaling this volume by ten times to 100M input tokens and 20M output tokens per month yields a revised bill of $34.90, compared to $36.40 under the old pricing. While this deepseek price cut generates tangible savings, the absolute dollar difference remains minimal for small to mid-sized API consumers running moderate traffic.

For massive operations running 1B input tokens and 200M output tokens per month, the old bill of $364.00 drops to $349.00, yielding $15.00 in monthly savings. At all scales, the proportional savings remain fixed at approximately 4.1 percent for this specific 5-to-1 input-to-output token ratio.

  • Old input bill: 10M input tokens * ($0.28 / 1M) = $2.80
  • Old output bill: 2M output tokens * ($0.42 / 1M) = $0.84
  • Total old monthly cost: $2.80 + $0.84 = $3.64
  • New input bill: 10M input tokens * ($0.269 / 1M) = $2.69
  • New output bill: 2M output tokens * ($0.40 / 1M) = $0.80
  • Total new monthly cost: $2.69 + $0.80 = $3.49

How does DeepSeek V3.2 Exp position against market median rates?

DeepSeek V3.2 Exp output rates sit far below the median output price recorded across tracked AI models. The median output rate across 36 tracked current models in our database is $3.90 per 1M tokens, compared to $0.40 per 1M output tokens for DeepSeek V3.2 Exp.

At $0.40 per 1M output tokens, DeepSeek V3.2 Exp costs less than one-ninth of the market median for generated text. Even before the 4.8 percent cut from $0.42, the model was positioned as an economical option relative to mainstream foundation models across commercial API providers.

This favorable output rate means that despite being pricier than specialized Flash options, DeepSeek V3.2 Exp remains an affordable alternative when compared against the broader commercial model catalog tracked by our monitors. For applications requiring capabilities beyond budget Flash tiers, V3.2 Exp maintains strong price positioning.

Who should keep using DeepSeek V3.2 Exp?

Teams with existing production applications tuned for DeepSeek V3.2 Exp should continue using the model if its capabilities match their workload requirements. The deepseek v3.2 exp price change delivers minor bill savings without requiring engineering teams to spend time re-testing prompts or modifying system architecture.

If your application relies on the 164k token context window and requires model behavior distinct from smaller Flash variants, staying on V3.2 Exp makes practical sense. The output price of $0.40 per 1M tokens provides strong economic value relative to the $3.90 per 1M token market median across tracked current models.

Organizations that run balanced workloads where output generation is moderate will find the price adjustment a welcome operational benefit that requires zero migration effort. Switching infrastructure to save pennies per million tokens is rarely worth the developer overhead required for refactoring.

Where does paying for DeepSeek V3.2 Exp stop making sense?

Paying for DeepSeek V3.2 Exp stops making sense when your workload can run effectively on high-speed Flash architectures. Migrating a 10M input and 2M output token workload to DeepSeek V4 Flash 0731 slashes monthly token expenses from $3.49 to $0.63, delivering an 81.9 percent cost savings.

Evaluating an llm pricing update requires comparing unit costs against cheaper alternatives that deliver high throughput. The math for a 10M input and 2M output monthly token workload highlights the financial gap across current catalog options:

DeepSeek V3.2 Exp costs $3.49 per month under the new rate card. DeepSeek V4 Flash 0731 costs $0.63 per month ($0.45 input plus $0.18 output). Qwen3.7 Flash costs $0.56 per month ($0.30 input plus $0.26 output). Nemotron 3.5 Lightning costs $1.20 per month ($0.80 input plus $0.40 output).

Migrating that workload from DeepSeek V3.2 Exp to DeepSeek V4 Flash 0731 saves $2.86 monthly. Furthermore, DeepSeek V4 Flash 0731 provides a 1311k token context window, the widest context in the current dataset, compared to the 164k context of V3.2 Exp. If your task quality does not degrade on Flash models, remaining on DeepSeek V3.2 Exp means paying a premium you do not need to spend.

  • DeepSeek V3.2 Exp (New): $3.49 per month
  • Nemotron 3.5 Lightning: $1.20 per month
  • DeepSeek V4 Flash 0731: $0.63 per month
  • Qwen3.7 Flash: $0.56 per month

Frequently asked questions

When did the DeepSeek V3.2 Exp price change occur?
Our automated pricing monitor recorded the rate change on 2026-08-29. Output pricing moved from $0.42 to $0.40 per 1M tokens, while input pricing moved from $0.28 to $0.269 per 1M tokens.
How much money does the DeepSeek V3.2 Exp price cut save on standard workloads?
For a workload processing 10M input tokens and 2M output tokens monthly, the cost decreases from $3.64 to $3.49. This represents an overall monthly bill reduction of $0.15, or approximately 4.1 percent.
How does DeepSeek V3.2 Exp output pricing compare to market averages?
At $0.40 per 1M output tokens, DeepSeek V3.2 Exp is priced significantly below the market baseline. The median output rate across 36 tracked current models in our database is $3.90 per 1M tokens.
What are the cheapest alternatives to DeepSeek V3.2 Exp?
DeepSeek V4 Flash 0731 costs $0.045 input and $0.09 output per 1M tokens with a 1311k token context window. Qwen3.7 Flash charges $0.03 input and $0.13 output per 1M tokens, while Nemotron 3.5 Lightning charges $0.08 input and $0.20 output per 1M tokens.
What is the context window size for DeepSeek V3.2 Exp?
DeepSeek V3.2 Exp features a 164k token context window. Applications requiring larger prompt buffers can look to DeepSeek V4 Flash 0731, which offers the widest context window in our current dataset at 1311k tokens.

About the author

AIOPLY Pricing Desk

Independent AI cost research, verified against live provider rates

Every figure in this article was checked against the provider's own pricing page before publication. Where AI assistance is used to draft a routine price report, a human editor verifies the numbers and signs it off.

Our editorial policy

Quality standards

Prices were checked against provider documentation on Sep 6, 2026. Rates change without notice, so confirm current figures with the provider before committing a budget. We publish list prices only and take no payment for placement.

Report a correction

Sources and further reading

More from the blog