Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, a…
| Rank | Model | Usage | Benchmark | Score | Model slug | Context | Summary |
|---|---|---|---|---|---|---|---|
| #1 |
Claude Fable 5.1
Anthropic
|
81.6 | coding | 81.6 | anthropic/claude-fable-5.1 | 1.0M | Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, a… |
| #2 |
GPT-5.6 Sol
OpenAI
|
78.3 | coding | 78.3 | openai/gpt-5.6-sol-20260709 | 1.1M | GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is p… |
| #3 |
Claude Opus 5 (Adaptive Reasoning, Max Effort)
Claude Opus 5
|
78 | coding | 78 | anthropic/claude-opus-5 | 1.0M | Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at… |
| #4 |
GPT-6 Astra
OpenAI
|
76.9 | coding | 76.9 | openai/gpt-6-astra | 1.1M | GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep rese… |
| #5 |
Grok 4.6
SpaceXAI
|
76.8 | coding | 76.8 | x-ai/grok-4.6 | 500K | Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM. |
| #6 |
GPT-5.6 Terra
OpenAI
|
76.7 | coding | 76.7 | openai/gpt-5.6-terra | 1.1M | GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.… |
| #7 |
Claude Fable 5
Anthropic
|
76.5 | coding | 76.5 | anthropic/claude-5-fable-20260609 | 1.0M | Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file… |
| #8 |
Gemini 3.8 Flash
Google
|
76.3 | coding | 76.3 | google/gemini-3.8-flash | 1.0M | Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic task… |
| #9 |
Kimi K3
MoonshotAI
|
76.2 | coding | 76.2 | moonshotai/kimi-k3 | 1.0M | Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and… |
| #10 |
Gemini 3.7 Flash
Google
|
76.1 | coding | 76.1 | google/gemini-3.7-flash | 1.0M | Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed f… |
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is p…
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at…
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep rese…
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.…
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file…
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic task…
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and…
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed f…