AI API price history: every verified change, on the record

Providers rewrite their pricing pages and the old numbers vanish. This log keeps them: append-only, dated, and tied to official sources.

By the benchr team · · Each price is a dated rate-card snapshot, not a promise that the rate will continue · View changelog

Entries loggedsince June 1, 2026
Cheapest input / 1MDeepSeek V4-Flash, off-peak
Priciest output / 1MClaude Fable 5
Price spreadoutput, $0.66 to $50 per 1M

How this log works

The log follows three rules. Append-only: entries are not edited or removed after publication. Same-day verification: a number enters the log only on a day benchr confirms it on a provider pricing page or announcement, not from a search snippet or third-party table. Source attached: every entry links to the record used for verification.

The append-only log begins on June 1, 2026 and does not backfill earlier prices. Older documented changes appear below as separate context, with their own sources.

The log so far

Append-only entries from history.json: input/output per 1M tokens at time of logging
LoggedModelEventInputOutput
2026-09-03GPT-6 AstraNew model$10$50
2026-09-02Gemini 3.8 FlashNew model$0.75$3.75
2026-09-01Claude Fable 5.1New model$10$50
2026-08-31Claude Sonnet 5Price change$2$10
2026-08-28DeepSeek V4-ProPrice change$1.32$3.96
2026-08-28DeepSeek V4-FlashPrice change$0.44$1.32
2026-08-27Qwen3.8-FlashNew model$0.15$0.47
2026-08-24GPT-5.6 SolPrice change$4$20
2026-08-24Gemini 3.6 FlashPrice change$0.75$3.75
2026-08-21Kimi K3New model$3$15
2026-08-21Grok 4.6New model$2$6
2026-08-21GLM-5.3New model$1.4$4.4
2026-08-21Gemini 3.7 FlashNew model$0.75$3.75
2026-08-03Qwen3.8-MaxNew model$2$6
2026-07-13MiniMax M3Model added$0.3$1.2
2026-07-13Grok 4.5Model added$2$6
2026-07-13GPT-5.6 SolContext and pricing re-verified$5$30
2026-07-13GLM-5.2Model added$1.4$4.4
2026-07-09GPT-5.6New model$5$30
2026-07-08gpt-5-6Correction
2026-07-08gemini-3-5-proCorrection
2026-07-08claude-sonnet-5Correction
2026-07-03Claude Sonnet 5Price change$2$10
2026-07-01Claude Sonnet 5New model$4$20
2026-06-30Gemini 3.5 ProNew model$2.5$15
2026-06-10GPT-5.4New model$2.5$15
2026-06-10Claude Fable 5New model$10$50
2026-06-01Mistral Large 3New model$0.5$1.5
2026-06-01Llama 4 ScoutNew model
2026-06-01Llama 4 MaverickNew model
2026-06-01Kimi K2.6New model$0.95$4
2026-06-01Grok 4.3New model$1.25$2.5
2026-06-01GPT-5 MiniNew model$0.25$2
2026-06-01GPT-5.5New model$5$30
2026-06-01GPT-5New model$1.25$10
2026-06-01Gemini 3.5 FlashNew model$1.5$9
2026-06-01DeepSeek V4-ProNew model$0.435$0.87
2026-06-01DeepSeek V4-FlashNew model$0.14$0.28
2026-06-01Claude Opus 4.8New model$5$25

Most entries are baselines: the verified starting point each future change gets measured against. That's how an honest price database begins: you can't detect a change without a trustworthy "before."

Documented moves from before the log

These predate June 1, 2026, so they live outside the append-only dataset, but each is documented by the provider's own announcements and worth keeping in view:

The Opus lane fell two-thirds. Claude Opus 4 and 4.1 billed $15/$75 per million tokens through 2025. Since Opus 4.5 (November 2025), the lane bills $5/$25 — and both old models retire this summer, making the cut universal.

GPT-4o halved within months. The May 2024 launch snapshot billed $5/$15; later 2024 snapshots billed $2.50/$10. The original snapshot, still billing 2024 prices, shuts down October 23, 2026.

Reasoning got cheap. o1 launched December 2024 at $15/$60, and o1-pro hit $150/$600 in March 2025, OpenAI's priciest model ever. Today GPT-5.5 reasons better at $5/$30.

DeepSeek reset the recorded floor twice, in both directions. The June 1, 2026 log captured V4-Pro at $0.435/$0.87 after a reported 75% reduction on api-docs.deepseek.com. On August 16, 2026 DeepSeek raised the whole V4 line and split it into peak and off-peak halves, so V4-Pro reads $1.32/$3.96 peak and $0.66/$1.98 off-peak, and V4-Flash $0.44/$1.32 and $0.22/$0.66. Peak hours are 01:00–04:00 and 06:00–10:00 UTC, Monday to Friday. Those figures describe dated rate cards; they can change, so the current provider page remains the source for a new purchase or budget.

And the increases

Published base rates are only one source of cost changes. Google's October 16 cutoff moves Gemini 2.5 Flash traffic to a model billing 5× more per input token. xAI re-pointed legacy Grok API slugs to Grok 4.3 billing, changing the price of unchanged integrations. Gemini's over-200K-token surcharge also raised effective long-context rates for some workloads. For that reason, this log records retirements, routing changes, and migration effects alongside rate-card updates.

What gets logged

Six event types match the dataset schema: price_change, new_model, benchmark_update, deprecation_announced, sunset, and context_change. Pricing means published list rates for input, output, and, where offered, cached input. Promotional and negotiated enterprise rates stay out. Open-weight records with no license charge are logged as $0-license entries; that is not a $0 infrastructure claim. Verified changes also appear in the site changelog and affected pricing pages.

Frequently asked

Can I use this data in my own work?

Yes: CC BY 4.0. Articles, research, dashboards, commercial tools, all fine. Attribution is a link to benchr.org. The JSON downloads include every entry's official source.

Why no entries before June 1, 2026?

Entries are only added on the day they're verified against the provider's live page. Backfilled numbers from memory or stale snippets are how price datasets rot. Older documented moves appear as sourced context above, outside the log.

Do AI prices ever go up?

Rate cards rarely rise; effective costs do. Forced migrations (Gemini 2.5 → 3.5 Flash), slug re-pointing (Grok), and long-context surcharges all raised real bills in 2026. The log tracks those events, not just list prices.

How often is this updated?

Whenever a verified event lands, typically within a day or two of a provider announcement. The deprecations RSS feed and main feed carry the alerts.

Changelog

  • — Added four verified records for Gemini 3.7 Flash, Grok 4.6, GLM-5.3, and Kimi K3; regenerated every row and the displayed count from history.json.
  • — Reframed the DeepSeek V4-Pro reduction as a June 1 rate-card snapshot and removed the unsupported claim that the price would continue indefinitely.
  • — Page published. Log contains 14 entries (June 1 baselines plus the June 10 Fable 5 and GPT-5.4 additions). Pre-log context moves sourced to provider announcements.

Sources

  • benchr history.json — the append-only log (per-entry sources inside)
  • benchr deprecations.json — lifecycle events with official source URLs
  • Provider pricing pages: platform.claude.com · openai.com/api/pricing · ai.google.dev/gemini-api/docs/pricing · api-docs.deepseek.com · x.ai/api (all re-verified June 12, 2026)
  • Anthropic Claude Opus 4.5 announcement (Opus-lane repricing, Nov 2025) — anthropic.com/news/claude-opus-4-5