Compare AI Models Side by Side

Pick 2 to 4 models to compare benchmark scores, pricing, and specs in 2026.

Grok 3
Gemini 2.5 Pro
Qwen3.7 Max
Gemini 2.5 Flash
xAI
Grok 3

xAI · Grok

Google
Gemini 2.5 Pro

Google · Gemini

Alibaba
Qwen3.7 Max

Alibaba · Qwen

Google
Gemini 2.5 Flash

Google · Gemini

Specifications

Spec
Grok 3
Gemini 2.5 Pro
Qwen3.7 Max
Gemini 2.5 Flash
ProviderxAIGoogleAlibabaGoogle
FamilyGrokGeminiQwenGemini
Release date2025-02-172025-06-172026-04-012025-05-20
Open sourceNoNoYesNo
Context window131K tokens1.0M tokens1.0M tokens1.0M tokens
Max output------66K tokens
Input price / MTok$3.00$1.25*$1.25$0.30
Output price / MTok$15.00$10.00*$3.75$2.50
Speed--151 tok/s206 tok/s222 tok/s
Best forreasoning, codinglong-context, reasoning, visionreasoning, coding, long-contextbudget, coding, agents

* Gemini 2.5 Pro: $2.50/$15 per MTok for prompts over 200K tokens

Benchmark scores (% of max)

Scores normalized to percentage of each benchmark's maximum reported value. Only officially reported scores are included.

Cost per million tokens (USD)

Context window (thousands of tokens)

Output speed (tokens per second)

Notes

Grok 3: Trained with 10x compute of Grok 2 on the 200K-GPU Colossus cluster.
Gemini 2.5 Pro: Now considered legacy as Gemini 3.x is GA.
Qwen3.7 Max: Apache 2.0 open weights; available via Alibaba Cloud Model Studio and third-party APIs.
Gemini 2.5 Flash: First Flash model with thinking capabilities.
Back to all models

Related

Deciding which model to build on for a real product? Book a free consultation and we'll help you pick the right one for your workload and budget.

Grok 3 vs Gemini 2.5 Pro vs Qwen3.7 Max vs Gemini 2.5 Flash | AI Model Comparison