Models

The same prompt does not land the same way on every model, because the vendors document different preferences. These pages summarise what each one asks for, and link to the documentation it came from.

GPT-5.6Verified 2026-09-01

OpenAI

Rewards restraint. The current generation performs better with shorter, less repetitive prompts than its predecessors did.

3 models
ClaudeVerified 2026-09-01

Anthropic

The one family with a documented preference for XML structure, and the one that responds most to being told why an instruction matters.

4 models
GeminiVerified 2026-09-01

Google

Prefers short over persuasive. Gemini 3 answers direct instructions better than elaborate ones, and puts your question last when there is a lot of data.

4 models
Llama 4Verified 2026-09-01

Meta

Open weights, and historically the family that made self-hosting mainstream. Worth understanding — but Llama 4 has not been updated since May 2025, so treat it as a known quantity rather than a current recommendation.

2 models
Microsoft 365 CopilotVerified 2026-09-01

Microsoft

The only one of the five that can already see your files. That changes what a good prompt looks like: less structure, more pointing.

1 models
GrokVerified 2026-09-01

xAI

Very large context and a high reasoning ceiling — but almost no published prompting guidance for its text models, so treat advice about Grok with more caution than the rest.

3 models
Mistral 3Verified 2026-09-01

Mistral AI

The best prompting documentation of any provider here, an EU inference region, and almost the whole line-up available as open weights.

4 models
DeepSeek V4Verified 2026-09-01

DeepSeek

The cheapest option here by a wide margin and the only one that open-weights its actual frontier model — with the caveat that the hosted API stores data in China.

2 models

Compare all models

ProviderModelsContextMax outputGood for
OpenAIGPT-5.6 Sol
Flagship for complex professional work.
1.05M128KHard analysis, long documents, agentic workflows with many tool calls.
OpenAIGPT-5.6 Terra
Balances intelligence and cost.
1.05M128KThe sensible default for most drafting, summarising and reasoning work.
OpenAIGPT-5.6 Luna
Built for budget-conscious, high-volume work.
1.05M128KClassification, extraction, routing, and anything you run thousands of times.
AnthropicClaude Fable 5
Next-generation intelligence for long-running agents.
1M128KThe hardest reasoning and long-horizon agent work.
AnthropicClaude Opus 5
For complex agentic coding and enterprise work.
1M128KCoding, analysis, and multi-step professional work at moderate latency.
AnthropicClaude Sonnet 5
The best combination of speed and intelligence.
1M128KDay-to-day drafting, summarising and interactive use.
AnthropicClaude Haiku 4.5
The fastest model with near-frontier intelligence.
200K64KHigh-volume classification, extraction and quick turns.
GoogleGemini 3.7 Flash
Built for complex coding, agentic workflows and reliable multi-step execution.
Agentic and coding work where you want speed alongside multi-step reliability.
GoogleGemini 3.6 Flash
Balances speed and multimodal capability for general agentic tasks.
Mixed text-and-image work at interactive speed.
GoogleGemini 3.5 Flash-Lite
The fastest, most cost-effective model in its generation for high-throughput work.
Classification, tagging, and very high request volumes.
GoogleGemini 2.5 Pro
Previous-generation flagship for complex reasoning and coding.
Workloads already tuned against it, and cases where you want the older, more verbose default.
MetaLlama 4 Scout
17B active parameters across 16 experts (109B total), with a 10M-token context window.
10MVery long inputs — full codebases, archives, transcripts — on hardware you control.
MetaLlama 4 Maverick
17B active parameters across 128 experts (400B total).
General assistant and reasoning work where you want the stronger open model.
MicrosoftMicrosoft 365 Copilot
Assistant embedded across Word, Excel, Outlook, Teams and PowerPoint, grounded in your tenant’s content.
Summarising meetings and threads, drafting from your own documents, and finding things across the tenant.
xAIGrok 4.6
Frontier model for coding, agentic tasks and knowledge work.
500K[object Object]Hard reasoning and agentic work where you also want very long output.
xAIGrok 4.3
Efficient long-context alternative.
1MVery long documents at lower cost than the flagship.
xAIGrok 4.20
Available as reasoning, non-reasoning and multi-agent variants.
1MChoosing explicitly whether you want reasoning at all, which the frontier models do not allow.
Mistral AIMistral Large 3
State-of-the-art, open-weight, general-purpose multimodal model.
256KGeneral work where you want frontier-class output at a low price, or the option to self-host the same model.
Mistral AIMistral Medium 3.5
Frontier-class multimodal model for agentic and coding use.
256KAgentic and coding workloads, with documented strong adherence to system prompts.
Mistral AIMistral Small 4
Hybrid model unifying instruct, reasoning and coding in one.
256KThe value pick, and the one most people should start with.
Mistral AIMinistral 3 8B
Efficient small model with text and vision.
256KHigh volume, low cost, and running locally — it is about 5 GB quantised.
DeepSeekDeepSeek V4 Pro
1.6T total parameters with 49B active, the more capable tier.
1M384KHard work at a fraction of competitors’ prices, and very long generations.
DeepSeekDeepSeek V4 Flash
284B total parameters with 13B active, the fast tier.
1M384KHigh volume at the lowest price of anything on this site, with the same context and output limits as the larger model.

A dash means the vendor does not publish that figure in a comparable place. Open the provider page and follow its documentation link for current specifications.