induwara.lk
induwara.lkAI · Reference

LLM Price History Tracker — API $/token over time

The dated API price of every major large-language model — GPT, Claude, Gemini, Grok and hosted open-weights — from launch to today, in one table and trend chart. The cost time-machine shows what your monthly token workload would have cost in 2023, 2024, 2025 and now, and exactly how much you save today. 37 cited prices across 5 providers. Free, no signup.

By Induwara AshinsanaUpdated Jul 17, 2026
LLM price history & cost time-machine37 dated prices
Sources cited · as of 2026-07-17

Cost time-machine

M

e.g. prompts, context, RAG documents you send.

M

e.g. completions the model generates back.

Workload presets
Tier
Currency

Anthropic Claude balanced line: Claude 2.1 → Claude 3/3.5/3.7 Sonnet → Sonnet 4.

You pay 51.79% less today than at the end of 2023.

End of 2023
Claude 2.1
$560.00/mo
End of 2024
Claude 3.5 Sonnet
$270.00/mo51.79%
End of 2025
Claude Sonnet 4
$270.00/mo51.79%
Today
Claude Sonnet 4
$270.00/mo51.79%

Fastest-falling family: Google Gemini Flash — combined input+output price fell 64.29% from Gemini 1.5 Flash (≤128K) (2024-05-14) to Gemini 2.0 Flash (2025-02-05).

Price trend — Output price

Metric
$0$25$50$75$1002023202420252026
FrontierMidBudget

Full dated price list

ModelEffectiveInput $/1MOutput $/1MΔ vs prev
Claude Haiku 4.5
Anthropic · Budget
2025-10-15$1.00$5.00+25%
Grok-4
xAI · Frontier
2025-07-09$3.00$15.000%
Gemini 2.5 Pro (≤200K)
Google · Frontier
2025-06-17$1.25$10.00
Gemini 2.5 Flash
Google · Budget
2025-06-17$0.30$2.50
Claude Opus 4
Anthropic · Frontier
2025-05-22$15.00$75.000%
Claude Sonnet 4
Anthropic · Mid
2025-05-22$3.00$15.000%
GPT-4.1
OpenAI · Frontier
2025-04-14$2.00$8.00
GPT-4.1 mini
OpenAI · Mid
2025-04-14$0.40$1.60
GPT-4.1 nano
OpenAI · Budget
2025-04-14$0.10$0.40
Grok-3
xAI · Frontier
2025-04-09$3.00$15.000%
Claude 3.7 Sonnet
Anthropic · Mid
2025-02-24$3.00$15.000%
Gemini 2.0 Flash
Google · Budget
2025-02-05$0.10$0.40+33.33%
DeepSeek-R1 (official API)
Open-weights · Mid
2025-01-20$0.55$2.19+99.09%
DeepSeek-V3 (official API)
Open-weights · Budget
2024-12-26$0.27$1.10
Llama 3.3 70B (Together)
Open-weights · Mid
2024-12-06$0.88$0.88-2.22%
Claude 3.5 Haiku
Anthropic · Budget
2024-11-04$0.80$4.00+220%
Gemini 1.5 Flash (price cut)
Google · Budget
2024-10-01$0.075$0.30-71.43%
Gemini 1.5 Pro (price cut, ≤128K)
Google · Mid
2024-10-01$1.25$5.00-52.38%
Grok-2 (grok-beta)
xAI · Mid
2024-10-01$5.00$15.00
GPT-4o (2024-08-06)
OpenAI · Frontier
2024-08-06$2.50$10.00-33.33%
Llama 3.1 405B (Together)
Open-weights · Frontier
2024-07-23$3.50$3.50
GPT-4o mini
OpenAI · Budget
2024-07-18$0.15$0.60
Claude 3.5 Sonnet
Anthropic · Mid
2024-06-20$3.00$15.000%
Gemini 1.5 Flash (≤128K)
Google · Budget
2024-05-14$0.35$1.05
Gemini 1.5 Pro (≤128K)
Google · Mid
2024-05-14$3.50$10.50+600%
GPT-4o (2024-05-13)
OpenAI · Frontier
2024-05-13$5.00$15.00
Llama 3 70B (Together)
Open-weights · Mid
2024-04-18$0.90$0.90
Claude 3 Haiku
Anthropic · Budget
2024-03-04$0.25$1.25
Claude 3 Sonnet
Anthropic · Mid
2024-03-04$3.00$15.00
Claude 3 Opus
Anthropic · Frontier
2024-03-04$15.00$75.00
GPT-3.5 Turbo (0125)
OpenAI · Budget
2024-01-25$0.50$1.50-25%
Gemini 1.0 Pro
Google · Mid
2023-12-13$0.50$1.50
Claude 2.1
Anthropic · Mid
2023-11-21$8.00$24.00
GPT-3.5 Turbo (1106)
OpenAI · Budget
2023-11-06$1.00$2.000%
GPT-4 Turbo (1106-preview)
OpenAI · Frontier
2023-11-06$10.00$30.00
GPT-3.5 Turbo (0613)
OpenAI · Budget
2023-06-13$1.50$2.00
GPT-4 (8K)
OpenAI · Frontier
2023-03-14$30.00$60.00

All prices are standard public list rates in USD per 1,000,000 tokens, text generation only — excludes batch, cached-input, fine-tuning and region-specific SKUs. LKR display uses a fixed snapshot of Rs 305/USD (CBSL indicative), not a live rate. API prices change by provider announcement at any time; figures are accurate as of 2026-07-17— always confirm against each row's linked source before budgeting.

How it works

Every price on this page is a published, dated list figure taken from the provider's own pricing page or its launch/price-change announcement — never an estimate. Prices are stored in USD per 1,000,000 tokens exactly as published (where a provider quotes $/1K, it is normalised by ×1000, with no rounding of the source figure). The dataset is a curated reference: one representative row per model family per price change, text generation only. Batch, cached-input, fine-tuning and region-specific SKUs are excluded, and that limit is stated so coverage isn't overstated.

The cost time-machine is deterministic arithmetic. For a price point, the monthly cost of your workload is:

cost = inputTokensₘ × inputPer1M + outputTokensₘ × outputPer1M

Your token volumes are entered in millions and prices are per million, so the multiply is direct. This is cross-checked internally against the per-token identity — raw tokens × ($/1M ÷ 1,000,000) — which must return the same figure to the cent, so the engine is self-verifying.

To build the four dated snapshots, the tool picks — for your chosen tier — the representative model whose effective date is the latest on or before each boundary (end of 2023, 2024, 2025, and today), then costs your workload at each. Savings against a baseline year b are (cost_b − cost_today) ÷ cost_b, measured from the earliest priced snapshot; if both token inputs are zero every cost is zero and savings show as “—”.

The Δ vs previous column compares each row to the most recent earlier price in the same model family: Δ = (new − old) ÷ old, computed separately for input and output. A negative Δ (green) is a price cut; a positive Δ (red) is a rise. The blended metric weights input and output at 3:1 — a typical chat/RAG ratio — and is used only for the chart and table display, never for your actual time-machine cost. The optional LKR display multiplies USD by a fixed, cited snapshot of Rs 305/USD; it is not a live rate. Because list prices change by provider announcement at any time, the whole page is stamped with a LAST_VERIFIED date of 2026-07-17 and every row links to its source.

Worked examples

Frontier tier — 40M input + 10M output / month

  1. End-2023 snapshot = GPT-4 Turbo ($10 in / $30 out per 1M)
  2. Cost: 40 × $10 + 10 × $30 = $400 + $300 = $700/month
  3. Today snapshot = GPT-4.1 ($2 in / $8 out per 1M)
  4. Cost: 40 × $2 + 10 × $8 = $80 + $80 = $160/month
  5. Savings vs end-2023: (700 − 160) ÷ 700 = 77.1%
  6. In LKR today at Rs 305/USD: 160 × 305 = Rs 48,800/month

Budget tier — 100M input + 20M output / month

  1. End-2023 snapshot = GPT-3.5 Turbo ($1.50 in / $2 out per 1M)
  2. Cost: 100 × $1.50 + 20 × $2 = $150 + $40 = $190/month
  3. Today snapshot = GPT-4.1 nano ($0.10 in / $0.40 out per 1M)
  4. Cost: 100 × $0.10 + 20 × $0.40 = $10 + $8 = $18/month
  5. Savings vs end-2023: (190 − 18) ÷ 190 = 90.5%

Edge case — a price that went UP (Claude Haiku, Δ column)

  1. Claude 3 Haiku (Mar 2024): $0.25 in / $1.25 out per 1M
  2. Claude 3.5 Haiku (Nov 2024): $0.80 in / $4.00 out per 1M
  3. Δ input: (0.80 − 0.25) ÷ 0.25 = +220%
  4. Δ output: (4.00 − 1.25) ÷ 1.25 = +220%
  5. Shown in red — not every generation gets cheaper.

Frequently asked questions

Sources & references

Each price row links to its own source. The primary pricing pages and dated announcement archives used to pin figures are:

Prices were last cross-checked against these sources on 2026-07-17. As of that date the fastest-falling family tracked here is Google Gemini Flash, down 64.29% on combined input+output since its first listed price. API prices change by provider announcement at any time — always confirm against the linked source before committing to a budget.

Related tools

Rate this tool
Be the first to rate

Comments & feedback

Spotted a bug or want an improvement? Tell us — our team reviews every comment, and good ideas get built. Comments are public and anonymous.

Spotted a price that's changed, or a model I should add?

Email me at [email protected] — I update the dataset when providers announce changes.