AI Models Compared
Every major frontier model, grouped by provider - pricing, context window, benchmark scores, and what each one is actually best at. Continuously web-verified; last checked 2026-10-01.
Recently released
Pricing at a glance
| Model | Strong at | Context | Input / 1M | Output / 1M | Type |
|---|---|---|---|---|---|
| OpenAI (15 models) | |||||
| GPT-6.1 SolNew | Long context | 1,050,000 | $2.00 | $10.00 | Closed |
| GPT-6 SolNew | Cost efficiency | 1.05M | $2.00 (standard), $0.10 (cached input) | $10.00 | Closed |
| GPT-Image-2.5 FlareNew | Multimodal | 400000 | 8 | 30 | Closed |
| GPT-6 AstraNew | Coding | 1.05M | $10.00 | $50.00 | Closed |
| AstraNew | Reasoning | 1050000 | 10 | 50 | Closed |
| GPT-Live-1New | Speed | 128000 | Not applicable | Not applicable | Closed |
| GPT-5.6-CyberNewPreview | Coding | 1.05M | $12.50 | $75.00 | Closed |
| GPT-5.6 | Coding | 1.05M | 5 | 30 | Closed |
| GPT-5.6 SolPreview | Coding | 1050K | $5.00 | $30.00 | Closed |
| GPT-5.5 Instant | Speed | 1050000 | 5.00 | 30.00 | Closed |
| GPT-5.5 | Coding | 1,050,000 tokens (128,000 max output) | $5.00 | $30.00 | Closed |
| GPT-5.4 | Long context | 1,050,000 tokens (128,000 max output) | $2.50 | $15.00 | Closed |
| GPT-5 | Math | 400,000 tokens (128,000 max output) | $1.25 | $10.00 | Closed |
| o1 | Reasoning | 200,000 tokens (100,000 max output) | $15.00 | $60.00 | Closed |
| GPT-4o | Speed | 128,000 tokens (16,384 max output) | $2.50 | $10.00 | Closed |
| Anthropic (11 models) | |||||
| Claude Sonnet 5.5New | Speed | 1M | $2.00 | $10.00 | Closed |
| Claude Opus 5.5New | Coding | 1M | $4.00 | $20.00 | Closed |
| Claude Mythos 5.1NewPreview | Coding | 1M | $10.00 | $50.00 | Closed |
| Claude Fable 5.1New | - | 1M | $10.00 | $50.00 | Closed |
| Claude Opus 5 | Long context | 1M | $5.00 | $25.00 | Closed |
| Claude Sonnet 5 | Coding | 1M | $3.00 | $15.00 | Closed |
| Claude Fable 5 | Coding | Not changed (1M) | $10.00 | $0.25 (cache reads only; base output price $50 unchanged) | Closed |
| Claude Opus 4.8 | Coding | 1M | $5.00 | $25.00 | Closed |
| Claude Opus 4.7 | Coding | 1M | $5.00 | $25.00 | Closed |
| Claude Sonnet 4.6 | Long context | 1M | $3.00 | $15.00 | Closed |
| Claude Haiku 4.5 | Speed | 200K | $1.00 | $5.00 | Closed |
| Google (10 models) | |||||
| Gemini 4 ArgonNewPreview | Long context | 1M | $2.00 | $10.00 | Closed |
| Gemini 3.8 Flash TTSNew | Speed | 8192 tokens | $0.50 | $9.00 | Closed |
| Gemini 3.8 FlashNew | Coding | 1M | $0.75 | $3.75 | Closed |
| Gemini 3.5 TranscribeNew | Multimodal | 96000 | 2.50 | 12.00 | Closed |
| Gemini 3.7 FlashNew | - | 1M | $0.75 | $3.75 | Closed |
| Gemini 3.6 Flash | Speed | 1M | $1.50 | $7.50 | Closed |
| Gemini 3.5 Flash CyberPreview | Coding | 1000000 | 1.50 | 7.50 | Closed |
| Gemini Omni FlashPreview | Multimodal | 1048576 tokens | 1.50 | 17.50 | Closed |
| Nano Banana 2 Lite | Speed | 1M | 0.25 | 1.50 | Closed |
| Gemma 4 12B | Cost efficiency | 256K tokens | Free | Free | Open |
| Google DeepMind (6 models) | |||||
| Gemini 3.8 LiveNew | Speed | 128K | $3.00 | $12.00 | Closed |
| Gemini Robotics ER 2 | Multimodal | 128K | $2 | $10 | Closed |
| Gemini 3.5 | Coding | 1,048,576 tokens (Gemini 3.5 Flash; Pro variant not yet released) | $1.50 | $9.00 | Closed |
| Gemini 2.5 Pro | Long context | 1,048,576 tokens (1M) input; up to 65K output | $1.25 | $10.00 | Closed |
| Gemini 2.5 Flash | Cost efficiency | 1,048,576 tokens (1M) input; up to 65,535 output | $0.30 | $2.50 | Closed |
| Gemma 3 | Cost efficiency | 128K tokens (32K for the 1B variant) | Free | Free | Open |
| Alibaba (3 models) | |||||
| Qwen 3.8-Flash-NextNew | Cost efficiency | 262K | Free | Free | Open |
| Qwen 3.8 27BNew | Coding | 262K | Free | Free | Open |
| Qwen 3.8 MaxNew | Long context | 983616 | 2.00 | 6.00 | Closed |
| Cohere (3 models) | |||||
| Cohere Parse 5New | Cost efficiency | 8192 | 1.50 | Not announced | Closed |
| North Mini Code | Cost efficiency | 256K tokens | 0 | 0 | Closed |
| Command R+ | Long context | 128K tokens | $2.50 | $10.00 | Closed |
| DeepSeek (3 models) | |||||
| DeepSeek V4NewPreview | Long context | 1M tokens | Free | Free | Open |
| DeepSeek V4 FlashPreview | Long context | 1M tokens | Free | Free | Open |
| DeepSeek V3 | Cost efficiency | 128K tokens | Free | Free | Open |
| Meta (3 models) | |||||
| Muse GlimmerNew | Cost efficiency | 128K | Free | Free | Open |
| Muse Image | Cost efficiency | 65536 | $0.01/image | $0.01/image | Closed |
| Llama 4 | Long context | Up to 10M tokens (Scout); ~1M tokens (Maverick) | Free | Free | Open |
| Mistral AI (3 models) | |||||
| Robostral Navigate | Cost efficiency | Not announced | Not announced | Not announced | Closed |
| Mistral OCR 4 | Cost efficiency | 16K | $4.00 per 1,000 pages (API) / $2.00 per 1,000 pages (Batch) | $5.00 per 1,000 pages (Document AI annotation tier) | Closed |
| Mistral Large | Reasoning | 128000 | 2.00 | 6.00 | Closed |
| Alibaba (Qwen Team) (1 model) | |||||
| Qwen 3 | Cost efficiency | 128K tokens (32K for 0.6B/1.7B/4B dense variants) | Free | Free | Open |
| Amazon Web Services (1 model) | |||||
| Amazon Nova Pro | Multimodal | 300K tokens | $0.80 | $3.20 | Closed |
| Anonymous (Stealth) (1 model) | |||||
| Ox AlphaNewPreview | Long context | 1,048,576 (1M tokens) | $0.00 | $0.00 | Closed |
| Cisco (1 model) | |||||
| Antares-1B | Cost efficiency | 128000 | Free | Free | Open |
| Cognition (1 model) | |||||
| SWE-2New | Coding | 1M | 3.00 | 15.00 | Closed |
| Microsoft (1 model) | |||||
| MAI-Cyber-1-FlashPreview | Coding | 256000 | Not announced | Not announced | Closed |
| Moonshot AI (1 model) | |||||
| Kimi K3 | Coding | 1M | $3.00 | $15.00 | Open |
| Poolside (1 model) | |||||
| Laguna S 2.1 | Coding | 1M | Not announced | Not announced | Open |
| Salesforce (1 model) | |||||
| Salesforce KoaNewPreview | Reasoning | 1000000 | Not announced | Not announced | Open |
| SpaceXAI (1 model) | |||||
| Grok 4.6New | Cost efficiency | 500K | $2.00 | $6.00 | Closed |
| SpaceXAI (xAI) (1 model) | |||||
| Grok 4.5 | Cost efficiency | 500K | $2.00 | $6.00 | Closed |
| Springboards (1 model) | |||||
| FlintPreview | Speed | Not announced | Not announced | Not announced | Closed |
| Thomson Reuters (1 model) | |||||
| ThomsonNew | Cost efficiency | 262000 | Not announced | Not announced | Closed |
| TypeSafe AI (1 model) | |||||
| JevNewPreview | Cost efficiency | 32K | $0.042 | Free | Closed |
| Writer (1 model) | |||||
| Palmyra X6New | Cost efficiency | 1000000 | $2.00 | $8.00 | Closed |
| xAI (SpaceXAI) (1 model) | |||||
| Grok 4.7New | Coding | 500K | $2.00 | $6.00 | Closed |
| Z.ai (Zhipu AI) (1 model) | |||||
| GLM-5.3-FlashNew | Long context | 1M | $0.075 | $0.25 | Open |
| Zhipu AI / Z.ai (1 model) | |||||
| GLM-5.2 | Long context | 1M | $1.40 | $4.40 | Open |
Prices in USD. Open-source models are free to self-host; API pricing varies by provider.
OpenAI15 models
GPT-6.1 SolNew
1,050,000 ctx2026-09
Near-Astra intelligence for a fifth of the price
GPT-6 SolNew
1.05M ctx2026-09
Cost-efficient mid-tier model for complex professional work and coding
GPT-Image-2.5 FlareNew
400000 ctx2026-09
Faster, sharper, smarter image generation
GPT-6 AstraNew
1.05M ctx2026-09
The world's most intelligent and aligned model
AstraNew
1050000 ctx2026-09
OpenAI's next major model with critical cybersecurity capabilities
GPT-Live-1New
128000 ctx2026-09
A new generation of full-duplex voice models for natural human-AI conversation
GPT-5.6-CyberNew
1.05M ctx2026-08
Purpose-trained model for advanced cybersecurity research and vulnerability testing
GPT-5.6
1.05M ctx2026-07
OpenAI's next-generation GPT-5.6 model family: Sol, Terra, and Luna
GPT-5.6 Sol
1050K ctx2026-06
OpenAI's most capable and security-hardened frontier model, in limited preview
GPT-5.5 Instant
1050000 ctx2026-05
The fast, default ChatGPT model tuned for low-latency responses
GPT-5.5
1,050,000 tokens (128,000 max output) ctx2026-04
OpenAI's smartest general-purpose frontier model for professional work
GPT-5.4
1,050,000 tokens (128,000 max output) ctx2026-03
Capable, cost-efficient predecessor to GPT-5.5 with a 1M+ context window
GPT-5
400,000 tokens (128,000 max output) ctx2025-08
OpenAI's landmark August 2025 flagship: strong reasoning at a low price
o1
200,000 tokens (100,000 max output) ctx2024-12
OpenAI's first-generation deep-reasoning model that thinks before answering
GPT-4o
128,000 tokens (16,384 max output) ctx2024-05
OpenAI's versatile, fast multimodal workhorse (text + image)
Anthropic11 models
Claude Sonnet 5.5New
1M ctx2026-09
Faster, cheaper mid-tier model for everyday tasks and agentic work
Claude Opus 5.5New
1M ctx2026-09
The strongest-performing model we've tested to date
Claude Mythos 5.1New
1M ctx2026-09
Anthropic's frontier-tier model designed for advanced cybersecurity and critical infrastructure work
Claude Fable 5.1New
1M ctx2026-09
Point release of Anthropic's flagship, tuned for long-horizon agentic work with a large effective price cut.
Claude Opus 5
1M ctx2026-07
Thoughtful and proactive model matching Fable 5 intelligence at half the price
Claude Sonnet 5
1M ctx2026-06
The best combination of speed and intelligence, at near-Opus quality for a Sonnet price.
Claude Fable 5
Not changed (1M) ctx2026-06
Anthropic's most capable widely released model - frontier intelligence for long-running agents.
Claude Opus 4.8
1M ctx2026-05
Anthropic's top general-availability workhorse for complex agentic coding and enterprise work.
Claude Opus 4.7
1M ctx2026-04
The April 2026 Opus flagship - top-tier coding and vision, now superseded by Opus 4.8.
Claude Sonnet 4.6
1M ctx2026-02
Anthropic's most capable Sonnet-class model of early 2026, now superseded by Sonnet 5.
Claude Haiku 4.5
200K ctx2025-10
Anthropic's fastest, cheapest model with near-frontier intelligence.
Google10 models
Gemini 4 ArgonNew
1M ctx2026-09
Google's frontier model for real-world coding, enterprise knowledge work, and cybersecurity defense
Gemini 3.8 Flash TTSNew
8192 tokens ctx2026-09
Advanced text-to-speech with natural language voice design and line-by-line performance direction
Gemini 3.8 FlashNew
1M ctx2026-09
Most intelligent workhorse model with enhanced reasoning and coding
Gemini 3.5 TranscribeNew
96000 ctx2026-08
The most precise speech-to-text model yet
Gemini 3.7 FlashNew
1M ctx2026-08
Most intelligent workhorse model yet for coding and agents
Gemini 3.6 Flash
1M ctx2026-07
More token efficient and cheaper workhorse model for coding and knowledge work
Gemini 3.5 Flash Cyber
1000000 ctx2026-07
Cost-efficient cybersecurity-focused model for vulnerability detection and patching
Gemini Omni Flash
1048576 tokens ctx2026-06
Create anything from any input with conversational video editing
Nano Banana 2 Lite
1M ctx2026-06
The fastest, most cost-efficient image generation model in the Nano Banana family
Gemma 4 12B
256K tokens ctx2026-06
OpenGoogle's laptop-runnable open multimodal model with a unified encoder-free design.
Google DeepMind6 models
Gemini 3.8 LiveNew
128K ctx2026-09
Most advanced live dialogue model for natural, real-time voice conversations
Gemini Robotics ER 2
128K ctx2026-07
High-level reasoning brain for robots enabling video understanding, task orchestration, and multi-robot collaboration
Gemini 3.5
1,048,576 tokens (Gemini 3.5 Flash; Pro variant not yet released) ctx2026-05
Google's frontier model for agents and coding, made fast and cheap.
Gemini 2.5 Pro
1,048,576 tokens (1M) input; up to 65K output ctx2025-06
Google's advanced thinking model for complex reasoning, coding, and long context.
Gemini 2.5 Flash
1,048,576 tokens (1M) input; up to 65,535 output ctx2025-06
Google's price-performance workhorse with thinking and a 1M-token context.
Gemma 3
128K tokens (32K for the 1B variant) ctx2025-03
OpenGoogle's open, multimodal, multilingual long-context model family.
Alibaba3 models
Qwen 3.8-Flash-NextNew
262K ctx2026-08
OpenA 125B MoE architecture preview of Qwen4 with only 6B active parameters per token
Qwen 3.8 27BNew
262K ctx2026-08
OpenApache 2.0 dense multimodal model designed for local deployment on consumer hardware
Qwen 3.8 MaxNew
983616 ctx2026-08
2.4 trillion-parameter multimodal flagship model, claimed second only to Claude Fable 5
Cohere3 models
Cohere Parse 5New
8192 ctx2026-08
Cost-effective document parsing at enterprise scale
North Mini Code
256K tokens ctx2026-06
Cohere's first open-weight agentic coding model - 30B MoE, 3B active, runs on one H100.
Command R+
128K tokens ctx2024-08
Cohere's RAG- and tool-use-optimized model, still live but superseded by Command A.
DeepSeek3 models
DeepSeek V4New
1M tokens ctx2026-08
OpenOpen-weight 1.6T-param MoE frontier model with a 1M-token context built for agents.
DeepSeek V4 Flash
1M tokens ctx2026-07
OpenCompact 284B-param open MoE that keeps a 1M context at a fraction of Pro's cost.
DeepSeek V3
128K tokens ctx2024-12
OpenThe open-weight 671B-param MoE that put DeepSeek on the frontier map.
Meta3 models
Muse GlimmerNew
128K ctx2026-08
OpenOpen-weight agentic AI that runs on consumer GPUs and laptops
Muse Image
65536 ctx2026-07
Image generation model with visual reasoning capabilities built for Meta's ecosystem
Llama 4
Up to 10M tokens (Scout); ~1M tokens (Maverick) ctx2025-04
OpenMeta's natively multimodal open MoE herd with industry-leading context length.
Mistral AI3 models
Robostral Navigate
Not announced ctx2026-07
8B embodied navigation model for autonomous robot movement using single RGB camera and natural language instructions
Mistral OCR 4
16K ctx2026-06
State-of-the-art document OCR that turns PDFs and scans into structured Markdown.
Mistral Large
128000 ctx2024-02
Mistral's state-of-the-art, open-weight, general-purpose multimodal flagship.
Alibaba (Qwen Team)1 model
Amazon Web Services1 model
Anonymous (Stealth)1 model
Cisco1 model
Cognition1 model
Microsoft1 model
Moonshot AI1 model
Poolside1 model
Salesforce1 model
SpaceXAI1 model
SpaceXAI (xAI)1 model
Springboards1 model
Thomson Reuters1 model
TypeSafe AI1 model
Writer1 model
xAI (SpaceXAI)1 model
Z.ai (Zhipu AI)1 model
Zhipu AI / Z.ai1 model
Benchmark scores sourced from official provider release posts. Prices subject to change - check provider pricing pages for current rates.