Endpoint · evidenceGETapi.agentstools.dev /model/compare ?models=gpt-4o%2C+claude-sonnet-4-5%2C+gemini-2.5-flash
Compare two to five models side by side.
Read more
Compare two to five models side by side. Give a comma separated list of names or aliases and get each model normalized on price, context, capabilities, knowledge cutoff and coding benchmark, plus flags for the cheapest, the best benchmark and the largest context. Fused from open keyless catalogs with a citation on every model.Overview
Grade?
UNRATED
0 of 30 paid calls toward a letter
Price
$0.01 per call
Paid calls
none yet
delivery unknown until someone pays
More info
Seller?
URL
https://api.agentstools.dev/model/compare?models=gpt-4o%2C+claude-sonnet-4-5%2C+gemini-2.5-flash
Seen
first 34 days ago · last 18 days ago
45 free handshakes
Payment
Pays through
x402
Offered on?
Base USDC
Listed on?
Polygon, Arbitrum, Base, Solana
Disagrees with the offer
Score, from free handshakes onlyweights 2026.09.4
UN
RATED
RATED
0.818 liveness verified, no letter
To a letter · 30 paid calls
0 of 30
Held back by sample size, not by the seller · gap to 1.000 is 0.182: 0.100 uncertain ? + 0.082 short ? · 0.082 to A
| Component | Weight | Measured | Lower bound ? | Adds | Short ? | Uncertain ? | |
|---|---|---|---|---|---|---|---|
| latency p95costs the most it answered as fast as its class |
0.20 | 0.59 | 0.500 | 0.100 | −0.082 | −0.018 | |
| liveness it answered at all |
0.60 | 1.00 | 0.863 | 0.518 | 0 | −0.082 | |
| price stability the price stayed where it was listed |
0.20 | 1.00 | 1.000 | 0.200 | 0 | 0 | |
| Composite | 1.00 | 0.818 | gap to 1.000 = 0.182 · 0.082 short · 0.100 uncertain | 0.818 | −0.082 | −0.100 |
Not measured: correctness · nobody paid; honesty · nobody paid; schema conformance · no answer had a published shape to check
The multiplication, written out
0.200 × 0.500 latency p95
+ 0.600 × 0.863 liveness
+ 0.200 × 1.000 price stability
= 0.818 → UNRATED: 45 of 30 paid calls weights 2026.09.4 · 45 samples · seed cap_01M1A5A3B3VZDGFNN2R34A5TV0|2026-08-23T21:20:07.509Z|2026.09.4
Verify it yourself
npx teppi-check https://api.agentstools.dev/model/compare?models=gpt-4o%2C+claude-sonnet-4-5%2C+gemini-2.5-flashcurl -s https://api.teppi.xyz/v1/trust/cap_01M1A5A3B3VZDGFNN2R34A5TV0