Compare AI Models Side by Side

Pick 2 to 4 models to compare benchmark scores, pricing, and specs in 2026.

Qwen3.7 Max
Gemini 2.5 Pro
Grok 3
Gemini 2.5 Flash
Alibaba
Qwen3.7 Max

Alibaba · Qwen

Google
Gemini 2.5 Pro

Google · Gemini

xAI
Grok 3

xAI · Grok

Google
Gemini 2.5 Flash

Google · Gemini

Specifications

Spec
Qwen3.7 Max
Gemini 2.5 Pro
Grok 3
Gemini 2.5 Flash
ProviderAlibabaGooglexAIGoogle
FamilyQwenGeminiGrokGemini
Release date2026-04-012025-06-172025-02-172025-05-20
Open sourceYesNoNoNo
Context window1.0M tokens1.0M tokens131K tokens1.0M tokens
Max output------66K tokens
Input price / MTok$1.25$1.25*$3.00$0.30
Output price / MTok$3.75$10.00*$15.00$2.50
Speed206 tok/s151 tok/s--222 tok/s
Best forreasoning, coding, long-contextlong-context, reasoning, visionreasoning, codingbudget, coding, agents

* Gemini 2.5 Pro: $2.50/$15 per MTok for prompts over 200K tokens

Benchmark scores (% of max)

Scores normalized to percentage of each benchmark's maximum reported value. Only officially reported scores are included.

Cost per million tokens (USD)

Context window (thousands of tokens)

Output speed (tokens per second)

Notes

Qwen3.7 Max: Apache 2.0 open weights; available via Alibaba Cloud Model Studio and third-party APIs.
Gemini 2.5 Pro: Now considered legacy as Gemini 3.x is GA.
Grok 3: Trained with 10x compute of Grok 2 on the 200K-GPU Colossus cluster.
Gemini 2.5 Flash: First Flash model with thinking capabilities.
Back to all models

Related

Deciding which model to build on for a real product? Book a free consultation and we'll help you pick the right one for your workload and budget.

Qwen3.7 Max vs Gemini 2.5 Pro vs Grok 3 vs Gemini 2.5 Flash | AI Model Comparison