Endpoint · inferenceGETmodelprices.xyz /llm/cheapest/million-token-context
Cheapest million-token-context LLM leaderboard: every AI model with a 1,000,000-token context window or larger, ranked by inference cost per token — Gemini 3 Pro, Gemini 3 Flash…
Read more
Cheapest million-token-context LLM leaderboard: every AI model with a 1,000,000-token context window or larger, ranked by inference cost per token — Gemini 3 Pro, Gemini 3 Flash, Llama 4 Scout, GPT-5 long-context tiers and more. Input, output and cache USD per 1M tokens with exact context window and max output joined in. Answers 'what is the cheapest model that fits my whole corpus?' Refreshed hourly.Overview
Grade?
UNRATED
0 of 30 paid calls toward a letter
Price
$0.01 per call
Paid calls
none yet
delivery unknown until someone pays
More info
Seller?
URL
https://modelprices.xyz/llm/cheapest/million-token-context
Seen
first 35 days ago · last 16 hr ago
63 free handshakes
Payment
Pays through
x402
Offered on?
Base USDC
Listed on?
Base
Handshakescomputed 2 days ago
0.888 from 58 free handshakes over 30 days · UNRATED ?
| Component | Weight | Measured | Lower bound ? | Adds | Short ? | Uncertain ? | |
|---|---|---|---|---|---|---|---|
| livenesscosts the most it answered at all |
0.60 | 1.00 | 0.894 | 0.536 | 0 | −0.064 | |
| latency p95 it answered as fast as its class |
0.20 | 0.88 | 0.760 | 0.152 | −0.024 | −0.024 | |
| price stability the price stayed where it was listed |
0.20 | 1.00 | 1.000 | 0.200 | 0 | 0 | |
| Composite | 1.00 | 0.888 | gap to 1.000 = 0.112 · 0.024 short · 0.088 uncertain | 0.888 | −0.024 | −0.088 |
Not measured: correctness · nobody paid; honesty · nobody paid; schema conformance · no answer had a published shape to check
BURN_INFREE_TIER_ONLY
Verify it yourself
npx teppi-check https://modelprices.xyz/llm/cheapest/million-token-contextcurl -s https://api.teppi.xyz/v1/trust/cap_01M1A5AFTTTVEQCTHRDDSWF9M5