Top AI Models
Current best-in-class AI models by category
GPT-5.6 Sol
OpenAI flagship reasoning and coding model for the most complex tasks.
Jul 2026
GPT-5.6 Terra
OpenAI balanced GPT-5.6 model for strong intelligence at lower cost than Sol.
Jul 2026
Gemini 3.5
Google DeepMind latest Gemini family combining frontier intelligence with action.
Jul 2026
gpt-oss 120B
OpenAI open reasoning model designed to run locally or in private infrastructure.
Jul 2026
Claude Fable 5
Anthropic next-generation model for long-running agents and high-intelligence work.
Jun 2026
Claude Sonnet 5
Anthropic model balancing speed, intelligence, and agentic coding performance.
Jun 2026
Claude Opus 4.8
Anthropic model for complex agentic coding and enterprise reasoning workloads.
May 2026
Gemini 3.1 Pro
Google's latest flagship model. State-of-the-art reasoning, multimodal, and agentic capabilities. Released Feb 19, 2026.
Feb 2026
Claude Sonnet 4.6
Anthropic's default free/Pro model. Full upgrade in coding, agents, long-context. 1M token context (beta). Released Feb 2026.
Feb 2026
Qwen3.5
Alibaba's open-weight MoE flagship. 397B total / 17B active params, 512 experts, 256K context, 201 languages, native multimodal.
Feb 2026
Grok 4.20
xAI's multi-agent system with 4-agent collaboration framework. SuperGrok exclusive. Feb 2026.
Feb 2026
Claude Opus 4.6
Anthropic's most capable model. Outperforms GPT-5.2 on GDPval-AA by ~144 Elo. 1M token context (beta). Released Feb 2026.
Feb 2026
GPT-5.3 Codex
OpenAI's best-in-class agentic coding model, combining Codex + GPT-5 stacks. ~25% faster inference, top SWE-bench scores.
Jan 2026
GPT-5.2
First model to cross 90% on ARC-AGI-1. GPT-5.2 Thinking scores 52.9% on ARC-AGI-2. Advanced reasoning flagship.
Dec 2025
GPT-5
OpenAI's flagship multimodal model, now default in ChatGPT. Replaces GPT-4o with automatic integrated reasoning.
Jul 2025
DeepSeek V3
Open-source frontier model. Context window expanded to 1M+ tokens (Feb 2026). Strong reasoning at very low inference cost.
Dec 2024
GLM-5.2
Z.ai released GLM-5.3 on August 14, 2026. The model reuses the 743B GLM-5.2 base unchanged. Every reported gain comes from scaled post-training: more long-horizon task environments, more environment types, longer training. Terminal-Bench 3.0 moves from 4.6 to 28.3, and DeepSWE v1.1 from 46.2 to 66.9
Aug 2026
GPT-OSS
Mistral AI has released Shieldstral 1.0 3B, an open-weights, policy-adaptive multimodal safety classifier that frames content moderation as a single yes/no question instead of a fixed harm taxonomy. Operators supply the policy as a plain-language query at inference time and get back a calibrated saf
Aug 2026
GPT-5.6 Luna
ChatGPT introduces improved GPT-5.6 Sol with better accuracy and consistency, plus expanded access for free users and unlimited everyday chats with GPT-5.6 Luna.
Aug 2026
Claude Mythos 5
OpenAI moved GPT-5.6 to general availability on July 9, 2026, shipping three tiers instead of one model. Sol is $5/$30 per 1M tokens, Terra is $2.50/$15, and Luna is $1/$6. Sol sets the Artificial Analysis Coding Agent Index at 80, 2.8 points above Claude Fable 5, and reaches 62.6% on OSWorld 2.0 us
Jul 2026
Gemini 3 Flash
Google's fast frontier model, default in Gemini app. PhD-level reasoning at speed. Replaced Gemini 2.5 Flash.
Jan 2026
Mistral Large 3
Mistral's open-source flagship. 41B active / 675B total MoE, 256K context, text + image. Apache 2.0.
Dec 2025
Grok 4
xAI's flagship model with real-time X/web access, 130K+ context, native tool use. Most intelligent model per xAI.
Jul 2025
o3 Pro
OpenAI's extended-thinking reasoning model for complex math, science and coding tasks.
Apr 2025
Llama 4 Maverick
Meta's open-source MoE model. 17B active params / 128 experts, native multimodal (text + image). Apache 2.0.
Apr 2025