Live LLM pricing - every model, every provider, one table

Input, output and cached token prices across providers, verified against primary sources. No sign-up, no client-side data grids.

Data as of 2026-08-01 · 338 models tracked
338
models
50
providers
18
MCP passports
7 🔴
traps found

Top price movements

Price history tracking since 2026-07

We record every change from the moment we start watching a source. No confirmed price changes yet - this is where they will appear.

What connecting an MCP server really costs

Keys, free-tier limits, traps and the schema token tax - cost passports for popular MCP servers, every number with a proof and a date.

Open the passports
18
passports
7 🔴
traps found
6
run for $0

Cheapest models by input price

See full pricing table
ModelContextInput$ / 1MOutput$ / 1MCached$ / 1M
Ling-2.6-flashinclusionAI262K0.010.030.002
Granite 4.0 MicroIBM131K0.0170.112-
Mistral NemoMistral131K0.020.04-
Nex-N2-MiniNex AGI262K0.0250.10.0025
Llama 3.2 1B InstructMeta131K0.0270.201-
gpt-oss-20bOpenAI131K0.030.130.03
Nova Micro 1.0Amazon128K0.0350.14-
gpt-oss-120bOpenAI131K0.0370.17-

What would it cost you?

Pick a task and a monthly volume - we show the three cheapest models.

  1. Ling-2.6-flash$3.40 per month
  2. Mistral Nemo$5.60 per month
  3. Llama 3 8B Lunaris$9.40 per month

Estimated for 200000 requests/month at 800 input / 300 output tokens per request. Open full calculator