Skip to content
Marko Ma

Compare model API prices

Published input and output rates in USD per million tokens. Compare models and estimate monthly cost.

Public retail API pricing only. Enterprise and privately negotiated rates, plus token subscription plans, are not included.

74 pricing entries across 10 providers · See each entry for its check date and official source

Prices across generations

Release date vs. today’s API price.

Price metric

Input · USD / 1M tokensLog scale · 10× steps

Current API prices across model generationsRelease date on the horizontal axis. Today's USD per million tokens on a logarithmic vertical axis. Lines connect releases in a family, not historical prices. Use Inspect release below for exact rates and official sources.$0.1$1$10$1002023202420252026GPT-4GPT-3.5 Turbo · 2023-03-01 · $0.50/1MGPT-4 · 2023-03-14 · $30.00/1MGPT-4o · 2024-05-13 · $2.50/1MGPT-4.1 · 2025-04-14 · $2.00/1MGPT-5 · 2025-08-07 · $1.25/1MGPT-5.5 · 2026-04-23 · $5.00/1MGPT-5.6 Sol · 2026-06-26 · $4.00/1MGPT-6 Sol · 2026-09-22 · $2.00/1MGPT-6.1 Sol · 2026-09-29 · $2.00/1MGPT-6.1 Sol$2.00Claude Sonnet 4.5 · 2025-09-29 · $3.00/1MClaude Sonnet 4.6 · 2026-02-17 · $3.00/1MClaude Sonnet 5 · 2026-06-30 · $2.00/1MClaude Sonnet 5.5 · 2026-09-28 · $2.00/1MSonnet 5.5$2.00DeepSeek V4.1 Flash · 2026-09-10 · $0.30/1MV4.1 Flash$0.30

First public release date →

OpenAI · Sep 29, 2026 · $2.00 in / $10.00 out per 1M

PricingRelease
Rate conditions

Checked 2026-09-29. Standard rate for prompts up to 272K input tokens; long-context rate is $4 input / $15 output per MTok.

Current rates, not price history. DeepSeek Flash: one currently priced release. Capabilities differ.

Browse prices

Filter the entries, then choose up to four to compare.

74 of 74 pricing entries

  1. AnthropicText models

    Claude Fable 5.1

    claude-fable-5-1

    Input / 1M$10.00

    Output / 1M$50.00

    Conditions & sources

    Base uncached Claude API rate; 5-minute prompt-cache writes, batch jobs, and other tiers have different rates.

    First public launch: Checked
  2. AnthropicText models

    Claude Haiku 4.5

    claude-haiku-4-5

    Input / 1M$1.00

    Output / 1M$5.00

    Conditions & sources

    Base uncached API rate.

    First public launch: Checked
  3. AnthropicText models

    Claude Opus 4.5

    claude-opus-4-5

    Input / 1M$5.00

    Output / 1M$25.00

    Conditions & sources

    Available legacy model. Base uncached API rate. Alias serves claude-opus-4-5-20251101; November 24 is the public launch date, not the snapshot date.

    First public launch: Checked
  4. AnthropicText models

    Claude Opus 4.6

    claude-opus-4-6

    Input / 1M$5.00

    Output / 1M$25.00

    Conditions & sources

    Available legacy model. Global uncached API rate; the full 1M context window uses standard pricing. Caching, batch and US-only inference have different rates.

    First public launch: Checked
  5. AnthropicText models

    Claude Opus 4.7

    claude-opus-4-7

    Input / 1M$5.00

    Output / 1M$25.00

    Conditions & sources

    Available legacy model. Global uncached API rate. The newer tokenizer can use more tokens than Opus 4.6 and earlier; caching, batch and US-only inference have different rates.

    First public launch: Checked
  6. AnthropicText models

    Claude Opus 4.8

    claude-opus-4-8

    Input / 1M$5.00

    Output / 1M$25.00

    Conditions & sources

    Base uncached API rate.

    First public launch: Checked
  7. AnthropicText models

    Claude Opus 5

    claude-opus-5

    Input / 1M$5.00

    Output / 1M$25.00

    Conditions & sources

    Base uncached API rate.

    First public launch: Checked
  8. AnthropicText models

    Claude Opus 5.5

    claude-opus-5-5

    Input / 1M$4.00

    Output / 1M$20.00

    Conditions & sources

    Base uncached Claude API rate; cache, batch, and other tiers are excluded.

    First public launch: Checked
  9. AnthropicText models

    Claude Sonnet 4.5

    claude-sonnet-4-5

    Input / 1M$3.00

    Output / 1M$15.00

    Conditions & sources

    Base uncached API rate. Deprecated September 30, 2026; retirement scheduled for November 30, 2026. Check availability before adopting.

    First public launch: Checked
  10. AnthropicText models

    Claude Sonnet 4.6

    claude-sonnet-4-6

    Input / 1M$3.00

    Output / 1M$15.00

    Conditions & sources

    Base uncached API rate.

    First public launch: Checked
  11. AnthropicText models

    Claude Sonnet 5

    claude-sonnet-5

    Input / 1M$2.00

    Output / 1M$10.00

    Conditions & sources

    Base uncached API rate; introductory $2/$10 pricing is now permanent.

    First public launch: Checked
  12. AnthropicText models

    Claude Sonnet 5.5

    claude-sonnet-5-5

    Input / 1M$2.00

    Output / 1M$10.00

    Conditions & sources

    Base uncached Claude API rate; cache, batch, and other tiers are excluded.

    First public launch: Checked
  13. DeepSeekText models

    DeepSeek V4 Pro (0813)

    deepseek-v4-pro

    Input / 1M$1.32

    Output / 1M$3.96

    Conditions & sources

    Peak uncached input/output rate; off-peak rates are 50% lower. API slug: deepseek-v4-pro, currently serving DeepSeek-V4-Pro-0813. Date is the 0813 GA release; the April V4 Pro preview is not separately priced today.

    First public launch: Checked
  14. DeepSeekText models

    DeepSeek V4.1 Flash

    deepseek-flash

    Input / 1M$0.30

    Output / 1M$1.20

    Conditions & sources

    Peak uncached input/output rate; off-peak rates are 50% lower. API slug: deepseek-flash. Retired V4 Flash and V4 Flash Vision Exp aliases now serve V4.1 Flash, not their original models.

    First public launch: Checked
  15. GoogleText models

    Gemini 3.1 Pro

    gemini-3.1-pro-preview

    Input / 1M$2.00

    Output / 1M$12.00

    Conditions & sources

    Current rate applies to prompts up to 200K tokens; higher rates apply above 200K.

    First public launch: Checked
  16. GoogleText models

    Gemini 3.7 Flash

    gemini-3.7-flash

    Input / 1M$0.75

    Output / 1M$3.75

    Conditions & sources

    Intro rate through Dec 31, 2026; $1.50/$7.50 from Jan 1, 2027.

    First public launch: Checked
  17. GoogleText models

    Gemini 3.8 Flash

    gemini-3.8-flash

    Input / 1M$0.75

    Output / 1M$3.75

    Conditions & sources

    Intro rate through Dec 31, 2026; $1.50/$7.50 from Jan 1, 2027.

    First public launch: Checked
  18. MetaText models

    Muse Spark 1.1

    muse-spark-1.1

    Input / 1M$1.25

    Output / 1M$4.25

    Conditions & sources

    Standard API tier; cached input is $0.15/MTok.

    First public launch: Checked
  19. MetaText models

    Muse Spark 1.2

    muse-spark-1.2

    Input / 1M$1.25

    Output / 1M$4.25

    Conditions & sources

    Standard API tier; cached input is $0.15/MTok. First public launch date not verified.

    First public launch: Not verifiedChecked
  20. MetaText models

    Muse Spark 1.3

    muse-spark-1.3

    Input / 1M$1.25

    Output / 1M$4.25

    Conditions & sources

    Standard API tier; cached input is $0.15/MTok. Contributor variants cost less in exchange for permission to train on prompts and completions and are excluded.

    First public launch: Checked
  21. MiniMaxText models

    MiniMax M2.7

    MiniMax-M2.7

    Input / 1M$0.30

    Output / 1M$1.20

    Conditions & sources

    Standard pay-as-you-go API rate. Token Plan subscription pricing is excluded.

    First public launch: Checked
  22. MiniMaxText models

    MiniMax M2.7 Highspeed

    MiniMax-M2.7-highspeed

    Input / 1M$0.60

    Output / 1M$2.40

    Conditions & sources

    Highspeed pay-as-you-go API variant. Token Plan subscription pricing is excluded.

    First public launch: Checked
  23. MiniMaxText models

    MiniMax M3

    MiniMax-M3

    Input / 1M$0.30

    Output / 1M$1.20

    Conditions & sources

    Standard pay-as-you-go tier for prompts up to 512K input tokens, at the published permanent 50% discount. Above 512K: $0.60 input / $2.40 output per MTok. Priority tier costs more.

    First public launch: Checked
  24. MistralText models

    Mistral Large 3

    mistral-large-2512

    Input / 1M$0.50

    Output / 1M$1.50

    Conditions & sources

    Standard direct API rates per 1M tokens.

    First public launch: Checked
  25. MistralText models

    Mistral Small 4

    mistral-small-2603

    Input / 1M$0.15

    Output / 1M$0.60

    Conditions & sources

    Standard direct API rates per 1M tokens.

    First public launch: Checked
  26. Moonshot AIText models

    Kimi K2.6

    kimi-k2.6

    Input / 1M$0.95

    Output / 1M$4.00

    Conditions & sources

    Direct Kimi API pay-as-you-go rate; cache hits are $0.16/MTok.

    First public launch: Checked
  27. Moonshot AIText models

    Kimi K2.7 Code

    kimi-k2.7-code

    Input / 1M$0.95

    Output / 1M$4.00

    Conditions & sources

    Direct Kimi API pay-as-you-go rate; cache hits are $0.19/MTok. Highspeed variant is separately priced.

    First public launch: Checked
  28. Moonshot AIText models

    Kimi K3

    kimi-k3

    Input / 1M$3.00

    Output / 1M$15.00

    Conditions & sources

    Direct Kimi API pay-as-you-go rate; cache hits are $0.30/MTok. Kimi membership and token plans are excluded.

    First public launch: Checked
  29. OpenAIText models

    Chat latest (GPT-5.5 Instant)

    chat-latest

    Input / 1M$5.00

    Output / 1M$30.00

    Conditions & sources

    Specialized ChatGPT model rate; cached input is $0.50/MTok.

    First public launch: Checked
  30. OpenAIText models

    GPT-3.5 Turbo

    gpt-3.5-turbo

    Input / 1M$0.50

    Output / 1M$1.50

    Conditions & sources

    Legacy model; current GPT-3.5 Turbo alias pricing, not the original 0301 snapshot price.

    First public launch: Checked
  31. OpenAIText models

    GPT-4

    gpt-4

    Input / 1M$30.00

    Output / 1M$60.00

    Conditions & sources

    Legacy model; standard direct API token rates.

    First public launch: Checked
  32. OpenAIText models

    GPT-4.1

    gpt-4.1

    Input / 1M$2.00

    Output / 1M$8.00

    Conditions & sources

    Standard direct API token rates; cache and batch discounts excluded.

    First public launch: Checked
  33. OpenAIText models

    GPT-4o

    gpt-4o

    Input / 1M$2.50

    Output / 1M$10.00

    Conditions & sources

    Current GPT-4o alias rate; earlier dated snapshots can have different prices.

    First public launch: Checked
  34. OpenAITranscription

    GPT-4o mini Transcribe

    gpt-4o-mini-transcribe

    Input / 1M$1.25

    Output / 1M$5.00

    Conditions & sources

    Transcription-token rates; listed estimated cost is $0.003/minute. First API release date not verified.

    First public launch: Not verifiedChecked
  35. OpenAITranscription

    GPT-4o Transcribe

    gpt-4o-transcribe

    Input / 1M$2.50

    Output / 1M$10.00

    Conditions & sources

    Transcription-token rates; listed estimated cost is $0.006/minute. First API release date not verified.

    First public launch: Not verifiedChecked
  36. OpenAIText models

    GPT-5

    gpt-5

    Input / 1M$1.25

    Output / 1M$10.00

    Conditions & sources

    Current GPT-5 alias rate; availability and retirement may differ by snapshot.

    First public launch: Checked
  37. OpenAIText models

    GPT-5.3 Codex

    gpt-5.3-codex

    Input / 1M$1.75

    Output / 1M$14.00

    Conditions & sources

    Codex-category standard rate: $1.75/$14 per MTok; Fast mode is $3.50/$28 per MTok. Cached input is $0.175/MTok. Public direct API release date/access not verified.

    First public launch: Not verifiedChecked
  38. OpenAIText models

    GPT-5.5

    gpt-5.5

    Input / 1M$5.00

    Output / 1M$30.00

    Conditions & sources

    First public release April 23; API availability April 24. Standard rate up to 272K input tokens; longer prompts cost $10 input / $45 output per MTok.

    First public launch: Checked
  39. OpenAIText models

    GPT-5.6 Cyber (Daybreak Red)

    gpt-5.6-cyber

    Input / 1M$12.50

    Output / 1M$75.00

    Conditions & sources

    Standard short-context rate; long-context pricing is not listed. Restricted Daybreak Red access; alias: gpt-daybreak-red-latest.

    First public launch: Checked
  40. OpenAIText models

    GPT-5.6 Luna

    gpt-5.6-luna

    Input / 1M$0.20

    Output / 1M$1.20

    Conditions & sources

    Current standard short-context rate; long-context rate is $0.40 input / $1.80 output per MTok. OpenAI reduced its price in July 2026.

    First public launch: Checked
  41. OpenAIText models

    GPT-5.6 Sol

    gpt-5.6-sol

    Input / 1M$4.00

    Output / 1M$20.00

    Conditions & sources

    Current promotional short-context rate; OpenAI says it lasts at least through Nov 21, 2026. Long-context rate is $8 input / $30 output per MTok. Daybreak Blue alias: gpt-daybreak-blue-latest.

    First public launch: Checked
  42. OpenAIText models

    GPT-5.6 Terra

    gpt-5.6-terra

    Input / 1M$2.00

    Output / 1M$12.00

    Conditions & sources

    Current standard short-context rate; long-context rate is $4 input / $18 output per MTok. GPT-5.6 launch pricing was $2.50/$15.

    First public launch: Checked
  43. OpenAIText models

    GPT-6 Astra

    gpt-6-astra

    Input / 1M$10.00

    Output / 1M$50.00

    Conditions & sources

    Standard short-context rate; long-context rate is $20 input / $75 output per MTok. Fast mode costs 2× standard.

    First public launch: Checked
  44. OpenAIText models

    GPT-6 Luna

    gpt-6-luna

    Input / 1M$0.10

    Output / 1M$0.50

    Conditions & sources

    Standard rate for prompts up to 272K input tokens; long-context rate is $0.20 input / $0.75 output per MTok.

    First public launch: Checked
  45. OpenAIText models

    GPT-6 Sol

    gpt-6-sol

    Input / 1M$2.00

    Output / 1M$10.00

    Conditions & sources

    Standard rate for prompts up to 272K input tokens; long-context rate is $4 input / $15 output per MTok. GPT-6.1 Sol is the newer Sol model.

    First public launch: Checked
  46. OpenAIText models

    GPT-6.1 Sol

    gpt-6.1-sol

    Input / 1M$2.00

    Output / 1M$10.00

    Conditions & sources

    Standard rate for prompts up to 272K input tokens; long-context rate is $4 input / $15 output per MTok.

    First public launch: Checked
  47. OpenAIImage generation

    GPT-Image-2 — image tokens

    gpt-image-2

    Input / 1M$4.00

    Output / 1M$15.00

    Conditions & sources

    Image-token rates; text input is $2.50/MTok. Cached image input is $1/MTok. First API release date not verified.

    First public launch: Not verifiedChecked
  48. OpenAIImage generation

    GPT-Image-2 — text input tokens

    gpt-image-2

    Input / 1M$2.50

    Output / 1MNot listed

    Conditions & sources

    Text input is $2.50/MTok; generated image output is priced separately at $15/MTok image tokens. First API release date not verified.

    First public launch: Not verifiedChecked
  49. OpenAIImage generation

    GPT-Image-2.5 Flare — image tokens

    gpt-image-2.5-flare

    Input / 1M$8.00

    Output / 1M$30.00

    Conditions & sources

    Image-token rates; text input is $5/MTok. Cached image input is $2/MTok.

    First public launch: Checked
  50. OpenAIImage generation

    GPT-Image-2.5 Flare — text input tokens

    gpt-image-2.5-flare

    Input / 1M$5.00

    Output / 1MNot listed

    Conditions & sources

    Text input is $5/MTok; generated image output is priced separately at $30/MTok image tokens.

    First public launch: Checked
  51. OpenAIImage generation

    GPT-Image-2.5 Sunburst — image tokens

    gpt-image-2.5-sunburst

    Input / 1M$8.00

    Output / 1M$30.00

    Conditions & sources

    Image-token rates; text input is $5/MTok. Cached image input is $2/MTok.

    First public launch: Checked
  52. OpenAIImage generation

    GPT-Image-2.5 Sunburst — text input tokens

    gpt-image-2.5-sunburst

    Input / 1M$5.00

    Output / 1MNot listed

    Conditions & sources

    Text input is $5/MTok; generated image output is priced separately at $30/MTok image tokens.

    First public launch: Checked
  53. OpenAIRealtime

    GPT-Realtime 2.1 — audio tokens

    gpt-realtime-2.1

    Input / 1M$32.00

    Output / 1M$64.00

    Conditions & sources

    Audio-token rates; cached audio input is $0.40/MTok. First API release date not verified.

    First public launch: Not verifiedChecked
  54. OpenAIRealtime

    GPT-Realtime 2.1 — image input tokens

    gpt-realtime-2.1

    Input / 1M$5.00

    Output / 1MNot listed

    Conditions & sources

    Image input is $5/MTok; no image output-token price is listed. First API release date not verified.

    First public launch: Not verifiedChecked
  55. OpenAIRealtime

    GPT-Realtime 2.1 — text tokens

    gpt-realtime-2.1

    Input / 1M$4.00

    Output / 1M$24.00

    Conditions & sources

    Text-token rates; cached text input is $0.40/MTok. First API release date not verified.

    First public launch: Not verifiedChecked
  56. OpenAIRealtime

    GPT-Realtime 2.1 mini — audio tokens

    gpt-realtime-2.1-mini

    Input / 1M$10.00

    Output / 1M$20.00

    Conditions & sources

    Audio-token rates; cached audio input is $0.30/MTok. First API release date not verified.

    First public launch: Not verifiedChecked
  57. OpenAIRealtime

    GPT-Realtime 2.1 mini — image input tokens

    gpt-realtime-2.1-mini

    Input / 1M$0.80

    Output / 1MNot listed

    Conditions & sources

    Image input is $0.80/MTok; no image output-token price is listed. First API release date not verified.

    First public launch: Not verifiedChecked
  58. OpenAIRealtime

    GPT-Realtime 2.1 mini — text tokens

    gpt-realtime-2.1-mini

    Input / 1M$0.60

    Output / 1M$2.40

    Conditions & sources

    Text-token rates; cached text input is $0.06/MTok. First API release date not verified.

    First public launch: Not verifiedChecked
  59. OpenAIFine-tuning

    o4-mini fine-tuned — data sharing

    o4-mini-2025-04-16

    Input / 1M$2.00

    Output / 1M$8.00

    Conditions & sources

    Fine-tuned inference rate with data sharing; training is $100/hour. Existing-user service only while OpenAI winds down the fine-tuning platform; date not verified.

    First public launch: Not verifiedChecked
  60. OpenAIFine-tuning

    o4-mini fine-tuned — standard

    o4-mini-2025-04-16

    Input / 1M$4.00

    Output / 1M$16.00

    Conditions & sources

    Fine-tuned inference rate; training is $100/hour. Existing-user service only while OpenAI winds down the fine-tuning platform; date not verified.

    First public launch: Not verifiedChecked
  61. xAIText models

    Grok 4.20 Multi-agent

    grok-4.20-multi-agent-0309

    Input / 1M$1.25

    Output / 1M$2.50

    Conditions & sources

    Global API rate below 200K prompt tokens; at 200K+ the full request costs $2.50 input / $5 output per MTok.

    First public launch: Checked
  62. xAIText models

    Grok 4.20 Non-reasoning

    grok-4.20-0309-non-reasoning

    Input / 1M$1.25

    Output / 1M$2.50

    Conditions & sources

    Global API rate below 200K prompt tokens; at 200K+ the full request costs $2.50 input / $5 output per MTok.

    First public launch: Checked
  63. xAIText models

    Grok 4.20 Reasoning

    grok-4.20-0309-reasoning

    Input / 1M$1.25

    Output / 1M$2.50

    Conditions & sources

    Global API rate below 200K prompt tokens; at 200K+ the full request costs $2.50 input / $5 output per MTok.

    First public launch: Checked
  64. xAIText models

    Grok 4.3

    grok-4.3

    Input / 1M$1.25

    Output / 1M$2.50

    Conditions & sources

    Global API rate below 200K prompt tokens; at 200K+ the full request costs $2.50 input / $5 output per MTok. Launch date not verified.

    First public launch: Not verifiedChecked
  65. xAIText models

    Grok 4.5

    grok-4.5

    Input / 1M$2.00

    Output / 1M$6.00

    Conditions & sources

    Global API rate below 200K prompt tokens; at 200K+ the full request costs $4 input / $12 output per MTok.

    First public launch: Checked
  66. xAIText models

    Grok 4.6

    grok-4.6

    Input / 1M$2.00

    Output / 1M$6.00

    Conditions & sources

    Global API rate below 200K prompt tokens; at 200K+ the full request costs $4 input / $12 output per MTok. US regional endpoint adds 10%.

    First public launch: Checked
  67. xAIText models

    Grok 4.7

    grok-4.7

    Input / 1M$2.00

    Output / 1M$6.00

    Conditions & sources

    Global API rate for prompts below 200K tokens; at 200K+ the full request costs $4 input / $12 output per MTok. US regional endpoint adds 10%.

    First public launch: Checked
  68. Z.aiText models

    GLM-4.7

    GLM-4.7

    Input / 1M$0.60

    Output / 1M$2.20

    Conditions & sources

    Uncached pay-as-you-go API rate.

    First public launch: Checked
  69. Z.aiText models

    GLM-5

    GLM-5

    Input / 1M$1.00

    Output / 1M$3.20

    Conditions & sources

    Uncached pay-as-you-go API rate.

    First public launch: Checked
  70. Z.aiText models

    GLM-5.1

    GLM-5.1

    Input / 1M$1.40

    Output / 1M$4.40

    Conditions & sources

    Uncached pay-as-you-go API rate.

    First public launch: Checked
  71. Z.aiText models

    GLM-5.2

    GLM-5.2

    Input / 1M$1.40

    Output / 1M$4.40

    Conditions & sources

    Uncached pay-as-you-go API rate.

    First public launch: Checked
  72. Z.aiText models

    GLM-5.3

    GLM-5.3

    Input / 1M$1.40

    Output / 1M$4.40

    Conditions & sources

    Uncached pay-as-you-go API rate.

    First public launch: Checked
  73. Z.aiText models

    GLM-5.3-Flash

    GLM-5.3-Flash

    Input / 1M$0.15

    Output / 1M$0.50

    Conditions & sources

    Uncached pay-as-you-go API rate. Limited-time free cache storage and other discounts are excluded.

    First public launch: Checked
  74. Z.aiText models

    GLM-5.3-FlashX

    GLM-5.3-FlashX

    Input / 1M$0.37

    Output / 1M$1.25

    Conditions & sources

    Uncached pay-as-you-go API rate. Public launch date not verified.

    First public launch: Not verifiedChecked

Compare current prices

Input and output rates for entries you choose above.

0 of 4 selected

Choose up to four entries from the price list to see their current rates on the same scale.

Estimate monthly token cost

Use one pricing entry and your expected token volume.

Estimated monthly cost$0.23

Uses the listed standard rates for the selected token modality. Excludes cached-token prices, batch or priority rates, long-context tiers, training, tools, and time-based charges.

About these rates and sources

Rates are direct-provider API prices in USD per million tokens for the use case or token modality named in each entry. Where a provider publishes multiple tiers, the entry gives the listed standard rate and notes other conditions. A missing output rate means the provider does not list an output-token rate for that entry.

First public launch dates are shown only when verified from an official announcement, model page, or API release note. They do not represent historical prices. Each entry links to its pricing source and shows when it was last checked.

Official provider sources