AI Models

The underlying AI models that power our agents, each with different capabilities.

Auto

Automatic model routing through the Vercel AI Gateway. Picks the best available model for each request and falls back to alternates if the primary is overloaded or unavailable.

Auto

Power 3/5

Auto

Free Smart default Zero Data Retention

Picks a capable model for you and keeps going if a provider is overloaded.

Anthropic Claude

Anthropic is an AI safety company and public benefit corporation that builds the Claude family of large language models. Founded in 2021, Anthropic focuses on developing reliable, steerable AI designed to be helpful, harmless, and honest.

Claude Opus

Power 5/5

Claude Opus 5

Premium Advanced reasoning Zero Data Retention

Our most capable model — frontier-level reasoning for complex agentic work, long-horizon tasks, and enterprise projects

Context 1M
Max output 128K
Cutoff May 2026

Claude Sonnet

Power 4/5

Claude Sonnet 5

Premium Recommended Zero Data Retention

The best combination of speed and intelligence, with adaptive thinking and long context support

Context 1M
Max output 128K
Cutoff January 2026

Claude Haiku

Power 2/5

Claude Haiku 4.5

Free Fastest Zero Data Retention

Fast and intelligent model for quick tasks

Context 200K
Max output 64K
Cutoff February 2025

Google Gemini

Google DeepMind's Gemini models offer multimodal understanding with large context windows, strong reasoning, and efficient performance across tasks.

Gemini Pro

Power 3/5

Gemini 3.1 Pro

Premium Experimental Zero Data Retention

Google's most advanced Gemini 3.1 model with powerful agentic capabilities, multimodal understanding, and state-of-the-art reasoning.

Context 1.05M
Max output 65.5K
Cutoff January 2025

Gemini Flash

Power 2/5

Gemini 3.6 Flash

Free Fast general use Zero Data Retention

Google's workhorse Flash model — stronger agentic and multimodal performance than 3.5 Flash with reduced token usage on long-horizon coding and reasoning tasks, at a lower output price. Dynamic thinking is on by default.

Context 1.05M
Max output 65.5K

Mistral AI

Mistral AI builds efficient, open-weight language models. Known for strong multilingual support, fast inference, and competitive performance at lower cost.

Mistral Medium

Power 4/5

Mistral Medium 3.5

Free Advanced reasoning Zero Data Retention

Mistral's most capable model — strong complex reasoning and multimodal understanding. Ahead of Mistral Large 3 despite the smaller-sounding name.

Context 256K
Max output 4.1K

Mistral Large

Power 2/5

Mistral Large 3

Free Open-weight Zero Data Retention

Open-weight model with strong multilingual support, at a third the price of Mistral Medium 3.5 — a step below it on complex reasoning.

Context 256K
Max output 4.1K

Mistral Small

Power 1/5

Mistral Small 4

Free Zero Data Retention

Hybrid model optimized for general chat, coding, agentic tasks, and complex reasoning with text and image input support

Context 256K
Max output 4.1K

OpenAI GPT

OpenAI's GPT models deliver strong general-purpose intelligence with broad knowledge, creative writing, and code generation capabilities.

GPT Sol

Power 5/5

GPT 5.6 Sol

Premium Most capable No Zero Data Retention

OpenAI's new flagship — token-efficient with stronger frontend design judgment, programmatic tool calling, and max reasoning effort for the hardest tasks. Priced like Claude Opus; available on the Pro plan and above.

Context 1M
Max output 272K
Cutoff February 16, 2026

Earlier versions

GPT Terra

Power 3/5

GPT 5.6 Terra

Premium No Zero Data Retention

GPT-5.6 Terra delivers strong performance at a lower price than Sol, with a 1M context window and optional reasoning. Sonnet-equivalent OpenAI option for Basic+ users.

Context 1.05M
Max output 128K
Cutoff February 16, 2026

Earlier versions

GPT Luna

Power 1/5

GPT 5.6 Luna

Free Fast No Zero Data Retention

GPT 5.6 Luna is an efficient, high-volume version of GPT-5.6 — fast and cost-effective for everyday tasks.

Context 400K
Max output 128K
Cutoff February 16, 2026

GPT OSS 120B

Power 1/5

GPT OSS 120B

Free Zero Data Retention

OpenAI's open-weights 120B-parameter model.

Context 128K
Max output 4.1K
Cutoff June 2024

xAI

xAI's Grok models — frontier reasoning with very large context windows.

Grok

Power 3/5

Grok 4.5

Premium Advanced reasoning Zero Data Retention

xAI's newest flagship, taking the top spot on independent agentic tool-use benchmarks — though its 500K token context window trades size for reasoning strength versus Grok 4.3's 1M.

Context 500K
Max output 32.8K

Grok Fast

Power 1/5

Grok 4.1 Fast

Free 2M context Zero Data Retention

xAI's fast model with an unprecedented 2 million token context window.

Context 2M
Max output 4.1K

DeepSeek

DeepSeek's efficient open models with strong coding and math reasoning.

DeepSeek Pro

Power 3/5

DeepSeek V4 Pro

Premium 1M context Zero Data Retention

DeepSeek's flagship V4 reasoning model with strong coding and math capabilities and a 1 million token context window.

Context 1M
Max output 32.8K

DeepSeek Flash

Power 2/5

DeepSeek V4 Flash

Premium 1M context Zero Data Retention

DeepSeek's fast and efficient V4 model optimized for speed with a 1 million token context window.

Context 1M
Max output 16.4K

Moonshot AI

Moonshot AI's Kimi models — long-context reasoning.

Kimi

Power 3/5

Kimi K3

Premium 1M context Zero Data Retention

Moonshot AI's new flagship — a 2.8T-parameter model with native vision input and long-horizon agentic reasoning across a 1 million token context window.

Context 1.05M
Max output 32.8K

Earlier versions

Z-AI

Z-AI's GLM models — fast, multilingual general-purpose intelligence.

GLM Fast

Power 2/5

GLM-5.2 Fast

Premium Fast, 1M context Zero Data Retention

A lower-latency, higher-throughput serving tier of GLM-5.2 for faster responses at a higher price.

Context 1M
Max output 4.1K

GLM

Power 1/5

GLM-5.2

Free 1M context Zero Data Retention

Z-AI's general-purpose model with strong multilingual capabilities and a 1 million token context window.

Context 1M
Max output 4.1K