COMPARE

Compare models

Pick up to three routes. Every value is read from the same registry that prices your requests — pricing version 2026-08-24.1. Unverified fields stay unverified: comparison never fills gaps with estimates.

Comparing 2 of 3 possible routes.

Side-by-side comparison of auto, auto:cheap
Attributeautoauto:cheap
TypeIntent — Fleet selectsIntent — Fleet selects
Billing tierfrontierfast
Context window1M tokens200K tokens
Max output128K tokens64K tokens
Input / MTok$6.25$1.25
Output / MTok$31.25$6.25
Caller toolsRelayedRelayed
Hosted toolsNone — hosted_tools: []None — hosted_tools: []
StreamingRejected with 400, not bufferedRejected with 400, not buffered
Capabilitiesreasoning, coding, long context, structured outputfast, cheap, structured output
LatencyNot measuredNot measured
Live availabilityNot probedNot probed

Unverified fields stay unverified here. Latency and live availability are shown as not measured because this deployment performs no latency sampling and no health probe — comparison never fills a gap with an estimate.

How to read the tier

Tier is what you are billed at, not a quality ranking. An intent reserves at the most expensive tier it may resolve to and settles at the tier of the model that actually answered, so a frontier-ceiling intent frequently costs less than its reserve. See pricing.

What is not here

No benchmark scores, no latency percentiles, no uptime figures. Modavio does not measure them, and publishing a number nobody sampled would make this table less useful than leaving the row honest.