Endpoint · evidenceGETapi.agentstools.dev /model/compare ?models=gpt-4o%2C+claude-sonnet-4-5%2C+gemini-2.5-flash
Compare two to five models side by side.
Read more
Compare two to five models side by side. Give a comma separated list of names or aliases and get each model normalized on price, context, capabilities, knowledge cutoff and coding benchmark, plus flags for the cheapest, the best benchmark and the largest context. Fused from open keyless catalogs with a citation on every model.Overview
Grade?
UNRATED
0 of 30 paid calls toward a letter
Price
$0.01 per call
Paid calls
none yet
delivery unknown until someone pays
More info
Seller?
URL
https://api.agentstools.dev/model/compare?models=gpt-4o%2C+claude-sonnet-4-5%2C+gemini-2.5-flash
Seen
first 34 days ago · last 18 days ago
45 free handshakes
Payment
Pays through
x402
Offered on?
Base USDC
Listed on?
Polygon, Arbitrum, Base, Solana
Disagrees with the offer
Handshakescomputed 11 days ago
0.818 from 45 free handshakes over 30 days · UNRATED ?
| Component | Weight | Measured | Lower bound ? | Adds | Short ? | Uncertain ? | |
|---|---|---|---|---|---|---|---|
| latency p95costs the most it answered as fast as its class |
0.20 | 0.59 | 0.500 | 0.100 | −0.082 | −0.018 | |
| liveness it answered at all |
0.60 | 1.00 | 0.863 | 0.518 | 0 | −0.082 | |
| price stability the price stayed where it was listed |
0.20 | 1.00 | 1.000 | 0.200 | 0 | 0 | |
| Composite | 1.00 | 0.818 | gap to 1.000 = 0.182 · 0.082 short · 0.100 uncertain | 0.818 | −0.082 | −0.100 |
Not measured: correctness · nobody paid; honesty · nobody paid; schema conformance · no answer had a published shape to check
BURN_INFREE_TIER_ONLY
Verify it yourself
npx teppi-check https://api.agentstools.dev/model/compare?models=gpt-4o%2C+claude-sonnet-4-5%2C+gemini-2.5-flashcurl -s https://api.teppi.xyz/v1/trust/cap_01M1A5A3B3VZDGFNN2R34A5TV0