The underlying AI models that power our agents, each with different capabilities.
Automatic routing through the Vercel AI Gateway. Sends each request to Claude Sonnet 5, and serves that same model from a backup host if the primary one is overloaded or unavailable.
Power 4/5
Picks a capable model for you and keeps going if a provider is overloaded.
Anthropic is an AI safety company and public benefit corporation that builds the Claude family of large language models. Founded in 2021, Anthropic focuses on developing reliable, steerable AI designed to be helpful, harmless, and honest.
Power 5/5
Our most capable model — frontier-level reasoning for complex agentic work, long-horizon tasks, and enterprise projects
Power 4/5
The best combination of speed and intelligence, with adaptive thinking and long context support
Power 2/5
Fast and intelligent model for quick tasks
Google DeepMind's Gemini models offer multimodal understanding with large context windows, strong reasoning, and efficient performance across tasks.
Power 3/5
Google's most advanced Gemini 3.1 model with powerful agentic capabilities, multimodal understanding, and state-of-the-art reasoning.
Power 2/5
Google's workhorse Flash model — a fast follow to 3.6 Flash at half the token price, with the same 1M context window and 65,536-token output ceiling. Dynamic thinking is on by default.
Mistral AI builds efficient, open-weight language models. Known for strong multilingual support, fast inference, and competitive performance at lower cost.
Power 4/5
Mistral's most capable model — strong complex reasoning and multimodal understanding. Ahead of Mistral Large 3 despite the smaller-sounding name.
Power 2/5
Open-weight model with strong multilingual support, at a third the price of Mistral Medium 3.5 — a step below it on complex reasoning.
Power 1/5
Hybrid model optimized for general chat, coding, agentic tasks, and complex reasoning with text and image input support
OpenAI's GPT models deliver strong general-purpose intelligence with broad knowledge, creative writing, and code generation capabilities.
Power 5/5
OpenAI's new flagship — token-efficient with stronger frontend design judgment, programmatic tool calling, and max reasoning effort for the hardest tasks. Priced like Claude Opus; available on the Pro plan and above.
Power 3/5
GPT 5.6 Luna is an efficient, high-volume version of GPT-5.6 — fast and cost-effective for everyday tasks, now with the same 1M context window as the rest of the 5.6 generation.
Power 3/5
GPT-5.6 Terra delivers strong performance at a lower price than Sol, with a 1M context window and optional reasoning. Sonnet-equivalent OpenAI option for Basic+ users.
Power 2/5
OpenAI's open-weights 120B-parameter model.
xAI's Grok models — frontier reasoning with very large context windows.
Power 4/5
xAI's newest flagship — a fast follow to Grok 4.5, keeping the 500K token context window at the same token price.
Power 1/5
xAI's fast reasoning model with a 1 million token context window.
DeepSeek's efficient open models with strong coding and math reasoning.
Power 3/5
DeepSeek's flagship V4 reasoning model with strong coding and math capabilities and a 1 million token context window.
Power 2/5
DeepSeek's fast and efficient V4 model optimized for speed with a 1 million token context window.
Moonshot AI's Kimi models — long-context reasoning.
Z-AI's GLM models — fast, multilingual general-purpose intelligence.
Power 2/5
A lower-latency, higher-throughput serving tier of GLM-5.2 for faster responses at a higher price.
Power 1/5
Z-AI's general-purpose model with strong multilingual capabilities and a 1 million token context window.
Automatic routing through the Vercel AI Gateway. Sends each request to Claude Sonnet 5, and serves that same model from a backup host if the primary one is overloaded or unavailable.
Power 4/5
Picks a capable model for you and keeps going if a provider is overloaded.
Anthropic is an AI safety company and public benefit corporation that builds the Claude family of large language models. Founded in 2021, Anthropic focuses on developing reliable, steerable AI designed to be helpful, harmless, and honest.
Power 5/5
Our most capable model — frontier-level reasoning for complex agentic work, long-horizon tasks, and enterprise projects
Power 4/5
The best combination of speed and intelligence, with adaptive thinking and long context support
Power 2/5
Fast and intelligent model for quick tasks
Google DeepMind's Gemini models offer multimodal understanding with large context windows, strong reasoning, and efficient performance across tasks.
Power 3/5
Google's most advanced Gemini 3.1 model with powerful agentic capabilities, multimodal understanding, and state-of-the-art reasoning.
Power 2/5
Google's workhorse Flash model — a fast follow to 3.6 Flash at half the token price, with the same 1M context window and 65,536-token output ceiling. Dynamic thinking is on by default.
Mistral AI builds efficient, open-weight language models. Known for strong multilingual support, fast inference, and competitive performance at lower cost.
Power 4/5
Mistral's most capable model — strong complex reasoning and multimodal understanding. Ahead of Mistral Large 3 despite the smaller-sounding name.
Power 2/5
Open-weight model with strong multilingual support, at a third the price of Mistral Medium 3.5 — a step below it on complex reasoning.
Power 1/5
Hybrid model optimized for general chat, coding, agentic tasks, and complex reasoning with text and image input support
OpenAI's GPT models deliver strong general-purpose intelligence with broad knowledge, creative writing, and code generation capabilities.
Power 5/5
OpenAI's new flagship — token-efficient with stronger frontend design judgment, programmatic tool calling, and max reasoning effort for the hardest tasks. Priced like Claude Opus; available on the Pro plan and above.
Power 3/5
GPT 5.6 Luna is an efficient, high-volume version of GPT-5.6 — fast and cost-effective for everyday tasks, now with the same 1M context window as the rest of the 5.6 generation.
Power 3/5
GPT-5.6 Terra delivers strong performance at a lower price than Sol, with a 1M context window and optional reasoning. Sonnet-equivalent OpenAI option for Basic+ users.
Power 2/5
OpenAI's open-weights 120B-parameter model.
xAI's Grok models — frontier reasoning with very large context windows.
Power 4/5
xAI's newest flagship — a fast follow to Grok 4.5, keeping the 500K token context window at the same token price.
Power 1/5
xAI's fast reasoning model with a 1 million token context window.
DeepSeek's efficient open models with strong coding and math reasoning.
Power 3/5
DeepSeek's flagship V4 reasoning model with strong coding and math capabilities and a 1 million token context window.
Power 2/5
DeepSeek's fast and efficient V4 model optimized for speed with a 1 million token context window.
Moonshot AI's Kimi models — long-context reasoning.
Z-AI's GLM models — fast, multilingual general-purpose intelligence.
Power 2/5
A lower-latency, higher-throughput serving tier of GLM-5.2 for faster responses at a higher price.
Power 1/5
Z-AI's general-purpose model with strong multilingual capabilities and a 1 million token context window.