Brass

The price of GPT‑4‑level AI fell 682-fold in about three years.

In March 2023 the only model at that level was GPT-4 itself, at $37.50 per million tokens. By July 2026 it was Qwen3.7 Flash, at $0.055. The newest top models did not get cheaper the same way: OpenAI’s flagship still lists in the tens of dollars per million tokens.

Brass AI price index, updated Oct 1, 2026.

$0.01$0.10$1$10$1002023202420252026GPT-4-level workOpenAI flagship$0.01$0.10$1$10$1002023202420252026GPT-4-level workOpenAI flagship
Cheapest list price for GPT-4-level results on GPQA Diamond, and the launch price of OpenAI’s flagship, in US dollars per million tokens, three input tokens per output token, log scale. The full table is below.

What $25 buys

At list price, $25 bought about 667,000 tokens of GPT-4-level work in March 2023 and about 455,000,000 in July 2026. Newer models often write more tokens per answer, so the cost of one correct answer on this test fell about 70-fold over the same period rather than 682-fold. A question cost about $0.031 on GPT-4 and $0.00044 on Qwen3.7 Flash with reasoning off, in the per-question cost data built on Epoch AI’s GPQA Diamond runs.

How we measure

For what each model costs today, per million words and read from its maker’s pricing page every hour, see the market board.

Download the data

prices.csv and prices.json. Free to reuse under CC BY 4.0; credit “Brass AI price index, brass.adorellc.pro”. We add a row when a cheaper GPT-4-level model or a new OpenAI flagship ships.

The full table

US dollars per million tokens.

GPT-4-level work

DateModelInputOutputBlendedSources
2023-03-14GPT-4 (gpt-4-0314)$30.00$60.00$37.501 17
2023-11-06GPT-4 Turbo (gpt-4-1106-preview)$10.00$30.00$15.002 17
2024-02-26Mistral Large (mistral-large-2402)$8.00$24.00$12.003 17
2024-03-04Claude 3 Sonnet$3.00$15.00$6.004 17
2024-03-13Claude 3 Haiku$0.25$1.25$0.505 17
2024-07-18GPT-4o mini$0.15$0.60$0.26256 17
2024-08-12Gemini 1.5 Flash (price cut)$0.075$0.30$0.13137 17
2025-01-30Qwen Turbo (qwen-turbo-2024-11-01)$0.05$0.20$0.08758 17
2026-07-27Qwen3.7 Flash$0.03$0.13$0.0559 17

OpenAI flagship

DateModelInputOutputBlendedSources
2023-03-14GPT-4$30.00$60.00$37.501
2023-11-06GPT-4 Turbo$10.00$30.00$15.002
2024-05-13GPT-4o$5.00$15.00$7.5010
2025-02-27GPT-4.5 preview$75.00$150.00$93.7511
2025-08-07GPT-5$1.25$10.00$3.437512
2025-12-11GPT-5.2$1.75$14.00$4.812513
2026-03-05GPT-5.4$2.50$15.00$5.62514
2026-04-24GPT-5.5$5.00$30.00$11.2515
2026-09-03GPT-6 Astra$10.00$50.00$20.0016

Sources

  1. OpenAI pricing page, archived 2023-03-15 (GPT-4 8K context: $0.03/1K prompt, $0.06/1K completion)
  2. OpenAI DevDay announcement, archived 2023-11-06 ($0.01/1K input, $0.03/1K output)
  3. Mistral pricing page, archived 2024-02-26 (Mistral Large $8/$24 per 1M)
  4. Anthropic, Introducing the next generation of Claude (Mar 4, 2024; Sonnet available in the API that day)
  5. Anthropic, Claude 3 Haiku release post (Mar 13, 2024); price from the Mar 4, 2024 Claude 3 announcement
  6. OpenAI, GPT-4o mini: advancing cost-efficient intelligence (Jul 18, 2024), archived
  7. Google Developers Blog, Gemini 1.5 Flash price drop effective Aug 12, 2024 (prompts under 128K), archived
  8. Alibaba Cloud Model Studio models page (international), archived 2025-01-30 (qwen-turbo $0.00005/1K input, $0.0002/1K output)
  9. Alibaba Cloud Model Studio pricing, International (Singapore) deployment, prompts up to 32K tokens
  10. OpenAI API pricing page, archived 2024-05-13 (gpt-4o-2024-05-13 $5/$15)
  11. OpenAI model page, GPT-4.5 Preview (snapshot gpt-4.5-preview-2025-02-27)
  12. OpenAI model page, GPT-5 (snapshot gpt-5-2025-08-07)
  13. OpenAI model page, GPT-5.2 (snapshot gpt-5.2-2025-12-11)
  14. OpenAI model page, GPT-5.4 (snapshot gpt-5.4-2026-03-05; prompts up to 272K)
  15. OpenAI model page, GPT-5.5 (prompts up to 272K); API release Apr 24, 2026 per OpenAI changelog
  16. OpenAI model page, GPT-6 Astra (prompts up to 272K); released Sep 3, 2026 per OpenAI changelog
  17. Epoch AI benchmark data, GPQA Diamond scores (downloaded 2026-10-01)