AI Cost Intelligence, Verified Daily.
Skip to content

Pricing database

AI model API pricing database and cost comparison

Every model we track, priced in USD per 1M tokens, with context limits, cached input rates, and the date each record was last verified. Public list rates only, no batch or enterprise discounts.

Models
43
Providers
12
Verified
Aug 28

Lowest input rate

Qwen3.7 Flash

$0.030 / 1M in

Widest context window

Grok 4 Fast

2M tokens

Showing 43 of 43 models.

Qwen3.7 Flash

Qwen · Verified Aug 28, 2026

current
In
$0.030
Out
$0.130
Cached
$0.0060
Context
1M

GPT-5 Nano

OpenAI · Verified Aug 28, 2026

current
In
$0.050
Out
$0.400
Cached
$0.0050
Context
400K

DeepSeek V4 Flash 0731

DeepSeek · Verified Aug 28, 2026

current
In
$0.060
Out
$0.120
Cached
$0.012
Context
1.31072M

Qwen3 32B

Hugging Face · Verified Aug 28, 2026

current
In
$0.080
Out
$0.280
Cached
n/a
Context
131K

Nemotron 3.5 Lightning

NVIDIA · Verified Aug 28, 2026

current
In
$0.100
Out
$0.250
Cached
$0.050
Context
262K

Qwen3.8 Flash

Qwen · Verified Aug 28, 2026

current
In
$0.150
Out
$0.470
Cached
$0.016
Context
1M

GPT-5.6 Luna

OpenAI · Verified Aug 28, 2026

current
In
$0.200
Out
$1.20
Cached
$0.020
Context
1.05M

GPT-5.6 Luna Pro

OpenAI · Verified Aug 28, 2026

current
In
$0.200
Out
$1.20
Cached
$0.020
Context
1.05M

Grok 4 Fast

xAI · Verified Aug 28, 2026

legacy
In
$0.200
Out
$0.500
Cached
n/a
Context
2M

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)

Google · Verified Aug 28, 2026

current
In
$0.250
Out
$1.50
Cached
n/a
Context
66K

GPT-5 Mini

OpenAI · Verified Aug 28, 2026

current
In
$0.250
Out
$2.00
Cached
$0.025
Context
400K

DeepSeek V3.2 Exp

DeepSeek · Verified Aug 28, 2026

current
In
$0.270
Out
$0.410
Cached
n/a
Context
164K

Llama 4 Maverick

Meta · Verified Aug 28, 2026

legacy
In
$0.270
Out
$0.850
Cached
n/a
Context
1M

Gemini 2.5 Flash

Google · Verified Aug 28, 2026

current
In
$0.300
Out
$2.50
Cached
$0.030
Context
1.048576M

Gemini 3.5 Flash Lite

Google · Verified Aug 28, 2026

current
In
$0.300
Out
$2.50
Cached
$0.030
Context
1.048576M

Gemini 3.7 Flash

Google · Verified Aug 28, 2026

current
In
$0.375
Out
$1.88
Cached
$0.037
Context
1.048576M

Qwen3.8 27B

Qwen · Verified Aug 28, 2026

current
In
$0.425
Out
$2.55
Cached
$0.085
Context
1M

Qwen3 235B A22B

Qwen · Verified Aug 28, 2026

current
In
$0.455
Out
$1.82
Cached
n/a
Context
131K

Llama Nemotron Ultra 253B

NVIDIA · Verified Aug 28, 2026

legacy
In
$0.600
Out
$1.80
Cached
n/a
Context
128K

DeepSeek V4 Pro 0813

DeepSeek · Verified Aug 28, 2026

current
In
$0.660
Out
$1.98
Cached
$0.022
Context
1.048576M

Gemini 3.6 Flash

Google · Verified Aug 28, 2026

current
In
$0.750
Out
$3.75
Cached
$0.075
Context
1.048576M

Qwen3 Max

Qwen · Verified Aug 28, 2026

current
In
$0.780
Out
$3.90
Cached
$0.156
Context
262K

Claude Haiku 4.5

Anthropic · Verified Aug 28, 2026

current
In
$1.00
Out
$5.00
Cached
$0.100
Context
200K

Sonar

Perplexity · Verified Aug 28, 2026

current
In
$1.00
Out
$1.00
Cached
n/a
Context
127K

GPT-5

OpenAI · Verified Aug 28, 2026

current
In
$1.25
Out
$10.00
Cached
$0.125
Context
400K

Claude Sonnet 5

Anthropic · Verified Aug 28, 2026

current
In
$2.00
Out
$10.00
Cached
$0.200
Context
1M

Gemini 3 Pro

Google · Verified Aug 28, 2026

legacy
In
$2.00
Out
$12.00
Cached
$0.200
Context
1M

Mistral Large 3

Mistral · Verified Aug 28, 2026

legacy
In
$2.00
Out
$6.00
Cached
n/a
Context
256K

GPT-5.6 Sol Pro

OpenAI · Verified Aug 28, 2026

current
In
$2.00
Out
$10.00
Cached
$0.200
Context
1.05M

GPT-5.6 Terra Pro

OpenAI · Verified Aug 28, 2026

current
In
$2.00
Out
$12.00
Cached
$0.200
Context
1.05M

GPT-5.6 Terra

OpenAI · Verified Aug 28, 2026

current
In
$2.00
Out
$12.00
Cached
$0.200
Context
1.05M

GPT-5.6 Sol

OpenAI · Verified Aug 28, 2026

current
In
$2.00
Out
$10.00
Cached
$0.200
Context
1.05M

Qwen3.8 2.4T A95B

Qwen · Verified Aug 28, 2026

current
In
$2.00
Out
$6.00
Cached
$0.250
Context
1.048576M

Qwen3.8 Max

Qwen · Verified Aug 28, 2026

current
In
$2.00
Out
$6.00
Cached
$0.250
Context
1M

Grok 4.5

xAI · Verified Aug 28, 2026

current
In
$2.00
Out
$6.00
Cached
$0.300
Context
500K

Grok 4.6

xAI · Verified Aug 28, 2026

current
In
$2.00
Out
$6.00
Cached
$0.500
Context
500K

GPT-4o

OpenAI · Verified Aug 28, 2026

current
In
$2.50
Out
$10.00
Cached
$1.25
Context
128K

Claude Sonnet 4.5

Anthropic · Verified Aug 28, 2026

current
In
$3.00
Out
$15.00
Cached
$0.300
Context
1M

Sonar Pro

Perplexity · Verified Aug 28, 2026

current
In
$3.00
Out
$15.00
Cached
n/a
Context
200K

Grok 4.1

xAI · Verified Aug 28, 2026

legacy
In
$3.00
Out
$15.00
Cached
$0.750
Context
256K

Claude Opus 5

Anthropic · Verified Aug 28, 2026

current
In
$5.00
Out
$25.00
Cached
$0.500
Context
1M

Claude Opus 4.5

Anthropic · Verified Aug 28, 2026

current
In
$5.00
Out
$25.00
Cached
$0.500
Context
200K

Claude Opus 5 (Fast)

Anthropic · Verified Aug 28, 2026

current
In
$10.00
Out
$50.00
Cached
$1.00
Context
1M

Rates are published list prices in USD per 1M tokens and exclude batch, committed-spend, and enterprise agreements. Run your own numbers in the cost calculators or put models head to head on the compare page.

Providers in this dataset

Anthropic

6 models

Claude family, long context and dependable instruction following.

DeepSeek

3 models

Aggressively priced reasoning models with cache discounts.

Google

6 models

Gemini family, very large context windows and low-cost tiers.

Hugging Face

1 models

Inference providers routing to open-weight models.

Meta

1 models

Llama open-weight models served by many inference providers.

Mistral

1 models

European lab with efficient small and mid-size models.

NVIDIA

2 models

Nemotron models served through NVIDIA NIM endpoints.

OpenAI

10 models

GPT family models, strong general reasoning and tool use.

OpenRouter

0 models

Aggregator routing one API across many model providers.

Perplexity

2 models

Sonar models with built-in web search grounding.

Qwen

7 models

Alibaba's Qwen line, open weights and low-cost hosted tiers.

xAI

4 models

Grok family, fast general models with large context.

Pricing database questions

What does the AI model pricing database cover?

It lists every model AIOPLY tracks with input price, output price, cached input price per 1M tokens, context window, max output, release date, and the date each record was last verified.

How often is AI model pricing updated?

Records are reconciled daily against a live pricing feed. Each row shows its own last verified date so you can see exactly how fresh a number is before you use it in a budget.

Are enterprise or batch discounts included?

No. All figures are published provider list rates in USD per 1M tokens. Batch processing, committed spend, and negotiated enterprise agreements usually price lower than the rates shown here.