Hard, high-stakes work
Start with each company's flagship model and test it on your real task.
GPT-5.6 Sol · Claude Opus 5 · Grok 4.5
Explore AI
Pick the kind of work you have, compare a few sensible options, then use your own task results to decide what earns a place in your stack.
Start with each company's flagship model and test it on your real task.
GPT-5.6 Sol · Claude Opus 5 · Grok 4.5
Use the middle tier when you want strong results without flagship cost.
GPT-5.6 Terra · Claude Sonnet 5 · Grok 4.3
Choose a low-cost model for volume, drafts, routing, and sub-agents.
GPT-5.6 Luna · Gemini Flash-Lite · Grok Build
Two labels worth knowing
Working memory
It can keep up to one million tokens in a single request. That total includes your prompt, files, chat history, and the model’s reply.
Very roughly, 1M tokens is around 750,000 words. More memory helps with large codebases and long documents.
API price
Input is everything you send to the model. Output is everything it generates back. Each price is the cost per one million tokens.
$1.50 in · $7.50 out means generated text costs more than the text you send. These are API prices, not app subscription prices.
6 models to compare
Anthropic
Anthropic's most capable model for demanding agentic work.
1M tokens
How much it can handle at once
$5 in · $25 out
You send it · it writes back
Anthropic
Built for complex agentic coding and long-horizon enterprise work.
1M tokens
How much it can handle at once
$5 in · $25 out
You send it · it writes back
Anthropic
Anthropic's most capable model, for demanding reasoning and long-horizon agentic work.
1M tokens
How much it can handle at once
$10 in · $50 out
You send it · it writes back
Anthropic
A fast frontier model balancing intelligence, speed, and cost.
1M tokens
How much it can handle at once
$2 in · $10 out
You send it · it writes back
Anthropic
Anthropic's fastest option for high-volume, cost-sensitive work.
200K tokens
How much it can handle at once
$1 in · $5 out
You send it · it writes back
Anthropic
A strong balance of intelligence, speed, and practical cost.
1M tokens
How much it can handle at once
$3 in · $15 out
You send it · it writes back
3 models to compare
Google's intelligent, speed-focused model with search and grounding.
1M tokens
How much it can handle at once
$1.5 in · $7.5 out
You send it · it writes back
Google's cost-efficient model for high-volume agentic tasks.
1M tokens
How much it can handle at once
$0.25 in · $1.5 out
You send it · it writes back
A multipurpose reasoning model for coding and complex work.
1.05M tokens
How much it can handle at once
$1.25 in · $10 out
You send it · it writes back
3 models to compare
OpenAI
Frontier model for complex professional reasoning and coding.
1.05M tokens
How much it can handle at once
$5 in · $30 out
You send it · it writes back
OpenAI
Balanced frontier intelligence for cost-aware production work.
1.05M tokens
How much it can handle at once
$2.5 in · $15 out
You send it · it writes back
OpenAI
Cost-sensitive model for high-volume production workloads.
1.05M tokens
How much it can handle at once
$1 in · $6 out
You send it · it writes back
3 models to compare
xAI
xAI's flagship model for reasoning with a large context window.
500K tokens
How much it can handle at once
$2 in · $6 out
You send it · it writes back
xAI
A lower-cost reasoning model with a one-million-token context.
1M tokens
How much it can handle at once
$1.25 in · $2.5 out
You send it · it writes back
xAI
A coding-focused model for building and editing software.
256K tokens
How much it can handle at once
$1 in · $2 out
You send it · it writes back