Compare models

GPT-4o mini vs Claude Sonnet 5 vs Claude Opus 5

Every figure below carries the source that reported it and the moment it was observed. Values that failed validation are withheld rather than shown, and a dash means no source we track carries that field for that model.
GPT-4o mini vs Claude Sonnet 5 vs Claude Opus 5, compared across 7 tracked fields. Each value shows the source that reported it and when.
FieldGPT-4o miniClaude Sonnet 5Claude Opus 5
Context window
128,000Synced from sourceCurrent3h ago
1,000,000 (highest of the compared values)Synced from sourceCurrent3h ago
1,000,000 (highest of the compared values)Synced from sourceCurrent3h ago
Input, $/M tokens
0.15 (lowest of the compared values)Synced from sourceCurrent3h ago
2Synced from sourceCurrent3h ago
5Synced from sourceCurrent3h ago
Modality
textSeeded, unreviewedwritten 29d ago
textSeeded, unreviewedwritten 29d ago
textSeeded, unreviewedwritten 29d ago
Licence
proprietarySeeded, unreviewedwritten 29d ago
proprietarySeeded, unreviewedwritten 29d ago
proprietarySeeded, unreviewedwritten 29d ago
OpenRouter id
openai/gpt-4o-miniSynced from sourceCurrent3h ago
anthropic/claude-sonnet-5Synced from sourceCurrent3h ago
anthropic/claude-opus-5Synced from sourceCurrent3h ago
Output, $/M tokens
0.6Synced from sourceCurrent3h ago
10Synced from sourceCurrent3h ago
25Synced from sourceCurrent3h ago
Vendor
OpenAISeeded, unreviewedwritten 29d ago
AnthropicSeeded, unreviewedwritten 29d ago
AnthropicSeeded, unreviewedwritten 29d ago

Sources: Hugging Face Hub, OpenRouter model catalogue. Most recent observation 3h ago. Where a row is marked, ▲ is the highest of the values shown and ▼ the lowest — arithmetic on the figures above, not a ranking or a recommendation. Rows where neither extreme is meaningful are left unmarked rather than given a direction they do not have.

Change the comparison

Remove one, or open the directory to pick a different set. Up to 3 at a time.


Every figure on this site resolves to the source that reported it, the moment it was observed, and how it was measured. A value that fails validation is withheld rather than shown — withheld means we refused to publish it, not that it is zero.