Data

LLM API Pricing Tracker

Current API prices per 1 million tokens for every major large language model, collected from official provider pricing pages and tracked over time. When a provider moves a price, the change is recorded here with the before and after.

Last verified: 2026-09-04 · 77 models · 5 providers · Updated daily

OpenAI

ModelInput / 1MOutput / 1MContextNotes
o1-pro$150.00$600.00short
gpt-5.5-pro$30.00$180.00
gpt-4-0613$30.00$60.00
gpt-5.4-pro$30.00$180.00
gpt-5.2-pro$21.00$168.00
o3-pro$20.00$80.00short
o1$15.00$60.00short
gpt-5-pro$15.00$120.00short
gpt-5.6-cyber$12.5$75.00short
gpt-5.5-cyber$12.5$75.00short
gpt-6-astra$10.00$50.00
gpt-4-turbo-2024-04-09$10.00$30.00
gpt-4o-2024-05-13$5.00$15.00
chat-latest$5.00$30.00
gpt-5.5$5.00$30.00Short-context rate; long-context tier is $10 in / $45 out
gpt-5.6-sol$4.00$20.00shortShort-context rate (<272K tokens); long-context tier is $8 in / $30 out
gpt-4o$2.5$10.00short
gpt-5.4$2.5$15.00Short-context rate; long-context tier is $5 in / $22.50 out
gpt-5.6-terra$2.00$12.00
gpt-4.1$2.00$8.00
davinci-002$2.00$2.00
o3$2.00$8.00short
gpt-5.3-codex$1.75$14.00
gpt-5.2$1.75$14.00
gpt-3.5-turbo-instruct$1.5$2.00
gpt-5$1.25$10.00short
gpt-5.1$1.25$10.00short
gpt-5-search-api$1.25$10.00
o4-mini$1.1$4.4short
o3-mini$1.1$4.4short
gpt-3.5-turbo-1106$1.00$2.00
gpt-5.4-mini$0.75$4.5
gpt-3.5-turbo-0125$0.5$1.5
gpt-3.5-turbo$0.5$1.5short
gpt-4.1-mini$0.4$1.6
babbage-002$0.4$0.4
gpt-5-mini$0.25$2.00short
gpt-5.4-nano$0.2$1.25
gpt-5.6-luna$0.2$1.2
gpt-4o-mini$0.15$0.6short
gpt-4.1-nano$0.1$0.4
gpt-5-nano$0.05$0.4short

Anthropic

ModelInput / 1MOutput / 1MContextNotes
claude-mythos-5$10.00$50.001MLimited availability via Glasswing partnership
claude-mythos-5.1$10.00$50.001MLimited availability via Glasswing partnership
claude-fable-5$10.00$50.001M
claude-fable-5.1$10.00$50.001M
claude-opus-4.5$5.00$25.001M
claude-opus-4-6$5.00$25.001M
claude-opus-5$5.00$25.001M
claude-opus-4-7$5.00$25.001M
claude-opus-4-8$5.00$25.001M
claude-sonnet-4.5$3.00$15.001M
claude-sonnet-4-6$3.00$15.001M
claude-sonnet-5$2.00$10.001MIntroductory pricing through August 31, 2026; standard pricing $3/$15 thereafter
claude-haiku-4-5$1.00$5.00200K

Google

ModelInput / 1MOutput / 1MContextNotes
gemini-3.1-pro-preview$2.00$12.00Rate for prompts up to 200K tokens; above 200K is $4 in / $18 out
gemini-3.5-flash$1.5$9.00
gemini-2.5-pro$1.25$10.00Rate for prompts up to 200K tokens; above 200K is $2.50 in / $15.00 out
gemini-3.8-flash$0.75$3.75Promotional rate through December 31, 2026; standard pricing $1.50/$7.50 thereafter
gemini-3.6-flash$0.75$3.75Promotional rate through December 31, 2026; standard pricing $1.50/$7.50 thereafter
gemini-3.7-flash$0.75$3.75Promotional rate through December 31, 2026; standard pricing $1.50/$7.50 thereafter
gemini-3-flash-preview$0.5$3.00Text/image/video input rate; audio input is $1.00
gemini-2.5-flash$0.3$2.5Text/image/video input rate; batch/flex tier is 50% off
gemini-3.5-flash-lite$0.3$2.5Hybrid reasoning model
gemini-3.1-flash-lite$0.25$1.5Text/image/video input rate; audio input is $0.50
gemini-2.5-flash-lite$0.1$0.4Most cost-efficient option
gemini-2.0-flash-lite$0.075$0.3

xAI

ModelInput / 1MOutput / 1MContextNotes
grok-4.5$2.00$6.001MMost capable xAI option for general tasks
grok-4.6$2.00$6.00500KShort-context rate (<200k); long-context (≥200k) is $4 in / $12 out
grok-4.20-0309-reasoning$1.25$2.51M
grok-4.20-multi-agent-0309$1.25$2.51M
grok-4.3$1.25$2.51M
grok-4.20-0309-non-reasoning$1.25$2.51M
grok-build-0.1$1.00$2.00256K

DeepSeek

ModelInput / 1MOutput / 1MContextNotes
deepseek-v4-pro$0.66$1.981MStandard off-peak rate (cache-miss input); cache-hit input significantly discounted; peak hours 01:00-04:00 and 06:00-10:00 UTC double pricing
deepseek-v4-flash-vision-exp$0.22$0.661MVision model; images are converted into tokens based on dimensions and billed as input tokens; cache-miss off-peak rate shown; cache-hit input $0.007 per MTok; peak hours 01:00-04:00 and 06:00-10:00 UTC double pricing
deepseek-v4-flash$0.22$0.661MStandard off-peak rate (cache-miss input); cache-hit input significantly discounted; peak hours 01:00-04:00 and 06:00-10:00 UTC double pricing

Price Change Log

2026-08-23OpenAI gpt-5.6-solinput $5.00$4.00, output $30.00$20.00
2026-08-17DeepSeek deepseek-v4-flashinput $0.14$0.22, output $0.28$0.66
2026-08-17DeepSeek deepseek-v4-proinput $0.435$0.66, output $0.87$1.98
2026-08-15Google gemini-3.6-flashinput $1.5$0.75, output $7.5$3.75

Free JSON API

The full dataset is available as JSON, free to use with attribution and a link back to this page. CORS is enabled, so it works directly from the browser.

curl https://techfastforward.com/api/data/llm-pricing

Methodology

Prices are read daily from each provider's official pricing page (linked on every model row) and stored only when they differ from the last recorded value, so the change log reflects real price moves rather than re-checks. All figures are standard pay-as-you-go API rates in USD per 1 million tokens; batch, cached-input, and volume discounts are noted per model where they materially differ.

Coverage focuses on text-generation flagship and workhorse models from OpenAI, Anthropic, Google, xAI, and DeepSeek. Spot an error or a missing model? Email [email protected].

Maintained by Jordan Hale, TechFastForward. Read our latest pricing coverage in Big Tech and Model Releases.