Compare AI Models Side by Side
Pick 2 to 4 models to compare benchmark scores, pricing, and specs in 2026.
Claude Haiku 4.5
Anthropic · Claude
Claude Sonnet 4.6
Anthropic · Claude
o3
OpenAI · o-series
Gemini 2.5 Pro
Google · Gemini
Specifications
Benchmark scores (% of max)
Scores normalized to percentage of each benchmark's maximum reported value. Only officially reported scores are included.
Cost per million tokens (USD)
Context window (thousands of tokens)
Output speed (tokens per second)
Notes
Claude Haiku 4.5: Fastest Claude model; best for high-volume and real-time tasks.
o3: Full reasoning model; chain-of-thought reasoning tokens billed as output.
Gemini 2.5 Pro: Now considered legacy as Gemini 3.x is GA.
Related
Deciding which model to build on for a real product? Book a free consultation and we'll help you pick the right one for your workload and budget.