modelprices.xyz
x402Grade BLast checked 51 min ago
GET https://modelprices.xyz/llm/cheapest/million-token-context
Cheapest million-token-context LLM leaderboard: every AI model with a 1,000,000-token context window or larger, ranked by inference cost per token — Gemini 3 Pro, Gemini 3 Flash, Llama 4 Scout, GPT-5 long-context tiers and more. Input, output and cache USD per 1M tokens with exact context window and max output joined in. Answers 'what is the cheapest model that fits my whole corpus?' Refreshed hourly.
Price$0.01per model request · paid in USD on Base
Uptime100%
4/4 probes · 1226.5 ms
Score72Grade B
Paid usage, 30 days2payers · 2 calls
How to buy
This endpoint takes pay-per-call payments over x402. Call it once without paying to see its live price; any x402 client then pays and retries automatically.
curl -i -X GET "https://modelprices.xyz/llm/cheapest/million-token-context"
# HTTP/1.1 402 Payment Required: price, asset and payTo are in the responseRecent checks
| Checked (UTC) | Result | Status | Time | Live price |
|---|---|---|---|---|
| 2026-10-11 10:31 | Valid 402 | 402 | 1192 ms | $0.01 |
| 2026-10-11 06:15 | Valid 402 | 402 | 1261 ms | $0.01 |
| 2026-10-11 01:45 | Valid 402 | 402 | 1061 ms | $0.01 |
| 2026-10-10 21:30 | Valid 402 | 402 | 1682 ms | $0.01 |
Receipts
No test purchases yet. Graded test purchases are next on our list; until then every number here comes from unpaid checks.
What a check proves