LLM API Prices, Tracked

Every model's input, output and cache rates in one place โ€” with what changes next. Prices double on Jan 1, 2027 for the Gemini 3.x family.

Model Pricing Table

ModelInput
$/1M
Output
$/1M
Upcoming changeContext
Gemini 3.8 Flash
gemini-3.8-flash
$0.75 $3.75 $1.5 / $7.5 1M
Gemini 3.7 Flash
gemini-3.7-flash
$0.375 $1.875 $0.75 / $3.75 1M
Gemini 3.6 Flash
gemini-3.6-flash
$0.375 $1.875 $0.75 / $3.75 1M
Claude Opus 5
claude-opus-5
$5 $25 โ€” โ€”
Claude Sonnet 5
claude-sonnet-5
$2 $10 โ€” โ€”
Claude Opus 4
claude-opus-4
$15 $75 โ€” โ€”
Claude Sonnet 4
claude-sonnet-4
$3 $15 โ€” โ€”

Output prices include thinking tokens where applicable.

Why a Price Tracker?

LLM prices change constantly โ€” promotional launch rates expire, vendors cut prices overnight, and billing models (cached input, thinking tokens, audio) hide the real cost. This site tracks the official numbers and flags upcoming changes, so you can budget before your bill surprises you.

Frequently Asked Questions

Where does the data come from?

Official provider pricing pages (e.g. Google AI docs), captured and dated. Each model page links its source.

How often is it updated?

Automated checks run daily against provider pages; manual verification on every price change.

Do you cover subscription plans?

API pricing first; consumer subscription comparisons are on the roadmap.

Models