Grok 4
xAI's frontier reasoning model trained on the expanded Colossus cluster
Verdict
A real step up from Grok 3 on reasoning benchmarks, with native tool use and real-time knowledge via X integration. Massive compute budget keeps xAI competitive with the other frontier labs. Good alternative for teams wanting model diversity outside the OpenAI/Anthropic/Google trio.
Other Text Generation & Reasoning
- Claude 3.5 SonnetStable
Anthropic's older workhorse with 200K context and computer use
- Claude 4 SonnetProduction
Anthropic's previous-gen workhorse — still solid, no longer the default
- Claude Fable 5Production
Anthropic's most intelligent generally available model — Mythos-class reasoning
- Claude Sonnet 5Production
Anthropic's production workhorse — Fable-class reasoning at Sonnet latency and cost
- Command R+Stable
Cohere's enterprise model optimised for RAG and tool use
- DeepSeek R1Production
Open-weight chain-of-thought reasoning rivalling GPT-o1

