Data
LLM API Pricing Tracker
Current API prices per 1 million tokens for every major large language model, collected from official provider pricing pages and tracked over time. When a provider moves a price, the change is recorded here with the before and after.
Last verified: 2026-09-04 · 77 models · 5 providers · Updated daily
OpenAI
| Model | Input / 1M | Output / 1M | Context | Notes |
|---|---|---|---|---|
| o1-pro | $150.00 | $600.00 | short | |
| gpt-5.5-pro | $30.00 | $180.00 | — | |
| gpt-4-0613 | $30.00 | $60.00 | — | |
| gpt-5.4-pro | $30.00 | $180.00 | — | |
| gpt-5.2-pro | $21.00 | $168.00 | — | |
| o3-pro | $20.00 | $80.00 | short | |
| o1 | $15.00 | $60.00 | short | |
| gpt-5-pro | $15.00 | $120.00 | short | |
| gpt-5.6-cyber | $12.5 | $75.00 | short | |
| gpt-5.5-cyber | $12.5 | $75.00 | short | |
| gpt-6-astra | $10.00 | $50.00 | — | |
| gpt-4-turbo-2024-04-09 | $10.00 | $30.00 | — | |
| gpt-4o-2024-05-13 | $5.00 | $15.00 | — | |
| chat-latest | $5.00 | $30.00 | — | |
| gpt-5.5 | $5.00 | $30.00 | — | Short-context rate; long-context tier is $10 in / $45 out |
| gpt-5.6-sol | $4.00 | $20.00 | short | Short-context rate (<272K tokens); long-context tier is $8 in / $30 out |
| gpt-4o | $2.5 | $10.00 | short | |
| gpt-5.4 | $2.5 | $15.00 | — | Short-context rate; long-context tier is $5 in / $22.50 out |
| gpt-5.6-terra | $2.00 | $12.00 | — | |
| gpt-4.1 | $2.00 | $8.00 | — | |
| davinci-002 | $2.00 | $2.00 | — | |
| o3 | $2.00 | $8.00 | short | |
| gpt-5.3-codex | $1.75 | $14.00 | — | |
| gpt-5.2 | $1.75 | $14.00 | — | |
| gpt-3.5-turbo-instruct | $1.5 | $2.00 | — | |
| gpt-5 | $1.25 | $10.00 | short | |
| gpt-5.1 | $1.25 | $10.00 | short | |
| gpt-5-search-api | $1.25 | $10.00 | — | |
| o4-mini | $1.1 | $4.4 | short | |
| o3-mini | $1.1 | $4.4 | short | |
| gpt-3.5-turbo-1106 | $1.00 | $2.00 | — | |
| gpt-5.4-mini | $0.75 | $4.5 | — | |
| gpt-3.5-turbo-0125 | $0.5 | $1.5 | — | |
| gpt-3.5-turbo | $0.5 | $1.5 | short | |
| gpt-4.1-mini | $0.4 | $1.6 | — | |
| babbage-002 | $0.4 | $0.4 | — | |
| gpt-5-mini | $0.25 | $2.00 | short | |
| gpt-5.4-nano | $0.2 | $1.25 | — | |
| gpt-5.6-luna | $0.2 | $1.2 | — | |
| gpt-4o-mini | $0.15 | $0.6 | short | |
| gpt-4.1-nano | $0.1 | $0.4 | — | |
| gpt-5-nano | $0.05 | $0.4 | short |
Anthropic
| Model | Input / 1M | Output / 1M | Context | Notes |
|---|---|---|---|---|
| claude-mythos-5 | $10.00 | $50.00 | 1M | Limited availability via Glasswing partnership |
| claude-mythos-5.1 | $10.00 | $50.00 | 1M | Limited availability via Glasswing partnership |
| claude-fable-5 | $10.00 | $50.00 | 1M | |
| claude-fable-5.1 | $10.00 | $50.00 | 1M | |
| claude-opus-4.5 | $5.00 | $25.00 | 1M | |
| claude-opus-4-6 | $5.00 | $25.00 | 1M | |
| claude-opus-5 | $5.00 | $25.00 | 1M | |
| claude-opus-4-7 | $5.00 | $25.00 | 1M | |
| claude-opus-4-8 | $5.00 | $25.00 | 1M | |
| claude-sonnet-4.5 | $3.00 | $15.00 | 1M | |
| claude-sonnet-4-6 | $3.00 | $15.00 | 1M | |
| claude-sonnet-5 | $2.00 | $10.00 | 1M | Introductory pricing through August 31, 2026; standard pricing $3/$15 thereafter |
| claude-haiku-4-5 | $1.00 | $5.00 | 200K |
| Model | Input / 1M | Output / 1M | Context | Notes |
|---|---|---|---|---|
| gemini-3.1-pro-preview | $2.00 | $12.00 | — | Rate for prompts up to 200K tokens; above 200K is $4 in / $18 out |
| gemini-3.5-flash | $1.5 | $9.00 | — | |
| gemini-2.5-pro | $1.25 | $10.00 | — | Rate for prompts up to 200K tokens; above 200K is $2.50 in / $15.00 out |
| gemini-3.8-flash | $0.75 | $3.75 | — | Promotional rate through December 31, 2026; standard pricing $1.50/$7.50 thereafter |
| gemini-3.6-flash | $0.75 | $3.75 | — | Promotional rate through December 31, 2026; standard pricing $1.50/$7.50 thereafter |
| gemini-3.7-flash | $0.75 | $3.75 | — | Promotional rate through December 31, 2026; standard pricing $1.50/$7.50 thereafter |
| gemini-3-flash-preview | $0.5 | $3.00 | — | Text/image/video input rate; audio input is $1.00 |
| gemini-2.5-flash | $0.3 | $2.5 | — | Text/image/video input rate; batch/flex tier is 50% off |
| gemini-3.5-flash-lite | $0.3 | $2.5 | — | Hybrid reasoning model |
| gemini-3.1-flash-lite | $0.25 | $1.5 | — | Text/image/video input rate; audio input is $0.50 |
| gemini-2.5-flash-lite | $0.1 | $0.4 | — | Most cost-efficient option |
| gemini-2.0-flash-lite | $0.075 | $0.3 | — |
xAI
| Model | Input / 1M | Output / 1M | Context | Notes |
|---|---|---|---|---|
| grok-4.5 | $2.00 | $6.00 | 1M | Most capable xAI option for general tasks |
| grok-4.6 | $2.00 | $6.00 | 500K | Short-context rate (<200k); long-context (≥200k) is $4 in / $12 out |
| grok-4.20-0309-reasoning | $1.25 | $2.5 | 1M | |
| grok-4.20-multi-agent-0309 | $1.25 | $2.5 | 1M | |
| grok-4.3 | $1.25 | $2.5 | 1M | |
| grok-4.20-0309-non-reasoning | $1.25 | $2.5 | 1M | |
| grok-build-0.1 | $1.00 | $2.00 | 256K |
DeepSeek
| Model | Input / 1M | Output / 1M | Context | Notes |
|---|---|---|---|---|
| deepseek-v4-pro | $0.66 | $1.98 | 1M | Standard off-peak rate (cache-miss input); cache-hit input significantly discounted; peak hours 01:00-04:00 and 06:00-10:00 UTC double pricing |
| deepseek-v4-flash-vision-exp | $0.22 | $0.66 | 1M | Vision model; images are converted into tokens based on dimensions and billed as input tokens; cache-miss off-peak rate shown; cache-hit input $0.007 per MTok; peak hours 01:00-04:00 and 06:00-10:00 UTC double pricing |
| deepseek-v4-flash | $0.22 | $0.66 | 1M | Standard off-peak rate (cache-miss input); cache-hit input significantly discounted; peak hours 01:00-04:00 and 06:00-10:00 UTC double pricing |
Price Change Log
Free JSON API
The full dataset is available as JSON, free to use with attribution and a link back to this page. CORS is enabled, so it works directly from the browser.
curl https://techfastforward.com/api/data/llm-pricingMethodology
Prices are read daily from each provider's official pricing page (linked on every model row) and stored only when they differ from the last recorded value, so the change log reflects real price moves rather than re-checks. All figures are standard pay-as-you-go API rates in USD per 1 million tokens; batch, cached-input, and volume discounts are noted per model where they materially differ.
Coverage focuses on text-generation flagship and workhorse models from OpenAI, Anthropic, Google, xAI, and DeepSeek. Spot an error or a missing model? Email [email protected].
Maintained by Jordan Hale, TechFastForward. Read our latest pricing coverage in Big Tech and Model Releases.