Try every frontier AI model
40 models across chat, image, video, and music — one account, pay only for what you use.
Chat & reasoning
Claude Opus 5
Chat & reasoning
Anthropic's everyday flagship for complex agentic coding and enterprise work. Ranks first overall on the AA Intelligence Index v4.1 at 63 with max effort.
Try Claude Opus 5→Claude Fable 5
Chat & reasoning
Anthropic's highest-capability model for the hardest reasoning and long-horizon agentic work. Scores 62 on the AA Intelligence Index v4.1, just behind Opus 5 at max effort.
Try Claude Fable 5→Grok 4.6
NewChat & reasoning
xAI's frontier model for long-running agents, codebase work, and ambitious interactive or visual projects. Scores 61 on the AA Intelligence Index with 500K context and always-on reasoning.
Try Grok 4.6→GPT-5.6 Sol
Chat & reasoning
OpenAI's flagship. Scores 61 on the AA Intelligence Index v4.1 and sets the state of the art on the Coding Agent Index (80), with an `ultra` multi-agent mode for the hardest work.
Try GPT-5.6 Sol→Claude Opus 4.8
Chat & reasoning
Frontier all-rounder. Leads SWE-bench Pro (69%) at AA Intelligence Index v4.1 of 56 — the value pick just below Fable 5.
Try Claude Opus 4.8→Grok 4.5
Chat & reasoning
xAI's previous-generation coding and agentic model, trained alongside Cursor. Scores 56 on the AA Intelligence Index with 500K context, always-on reasoning, and optional web search.
Try Grok 4.5→Gemini 3.7 Flash
NewChat & reasoning
Google's most intelligent Flash workhorse for coding and agents. Scores 56 on the AA Intelligence Index with 1M context, multimodal input, and introductory half-price token rates through 2026.
Try Gemini 3.7 Flash→GPT-5.6 Terra
Chat & reasoning
Balanced GPT-5.6 tier — matches GPT-5.5 quality (AA Intelligence Index 55) at roughly half the cost, tuned for everyday professional work.
Try GPT-5.6 Terra→GPT-5.5
Chat & reasoning
OpenAI's previous flagship. AA Intelligence Index v4.1 of 54.8 — a strong all-round frontier model, now succeeded by the GPT-5.6 family.
Try GPT-5.5→Claude Sonnet 5
Chat & reasoning
Anthropic's current fast workhorse — near-Opus quality at much lower latency and cost.
Try Claude Sonnet 5→Gemini 3.6 Flash
Chat & reasoning
Google's previous Flash workhorse, scoring 52 on the AA Intelligence Index. Retained for its 1M context, multimodal input, and low-latency minimal reasoning mode.
Try Gemini 3.6 Flash→GPT-5.6 Luna
Chat & reasoning
Fastest, most affordable GPT-5.6 tier — outperforms Claude Opus 4.8 on the Coding Agent Index at roughly a quarter of the cost.
Try GPT-5.6 Luna→GLM-5.2
Chat & reasoning
Zhipu's GLM-5.2 — the leading open-weights model on the AA Intelligence Index.
Try GLM-5.2→Muse Spark 1.1
Chat & reasoning
Meta Superintelligence Labs' multimodal reasoning model, built for agentic work. 1M-token context, image + video understanding, search grounding, and strong coding — served on the new OpenAI-compatible Meta Model API at value pricing.
Try Muse Spark 1.1→Gemini 3.1 Pro
Chat & reasoning
Google DeepMind frontier. 2M context, fully multimodal, and an AA Intelligence Index v4.1 score of 46.
Try Gemini 3.1 Pro→MiniMax M3
Chat & reasoning
MiniMax M3 — Apache-licensed, matches Gemini 3.1 Pro on coding at open-weights pricing.
Try MiniMax M3→DeepSeek V4 Pro
Chat & reasoning
Flagship open-source MoE (1.6T) — frontier reasoning + coding with a 1M-token context, at open-weights pricing.
Try DeepSeek V4 Pro→Kimi K2.6
Chat & reasoning
Moonshot's K2.6 — open-weights, 80%+ SWE-bench Verified, top-tier agentic reasoning.
Try Kimi K2.6→Qwen 3.7 Plus
Chat & reasoning
Alibaba's flagship closed model — strong reasoning + function calling, served exclusively on Fireworks outside Alibaba's own cloud.
Try Qwen 3.7 Plus→Grok 4.3
Chat & reasoning
xAI flagship. Configurable reasoning, 1M context, and optional live web knowledge at low frontier pricing.
Try Grok 4.3→GPT-5
Chat & reasoning
OpenAI workhorse — strong general purpose at lower cost than 5.5.
Try GPT-5→Not sure which to pick? Read our guide to the best AI model for coding or compare models head-to-head on TryAI.