CanMyPCRunAILocal AI compatibility

Frontier models · snapshot 23 September 2026

Which AI model is best at what?

The most capable AI models right now, ranked on one independent benchmark and then split by task, because no single model wins everything. Every number links to where it came from.

Best at each job

Overall intelligence

Claude Opus 5.5

Highest Artificial Analysis Intelligence Index score at 58, five points clear of GPT-6 Astra and Claude Fable 5.1 (both 53).

Independent measurement

Agentic coding

Claude Opus 5.5

Terminal-Bench 4.0: 66.4% against 57.9% for GPT-6 Astra and 55.8% for Fable 5.1. Before Opus 5.5 arrived, Artificial Analysis had GPT-6 Astra and Fable 5.1 tied at 62 on its Coding Agent Index, with Astra doing it at about 60% of the cost.

Vendor-reported benchmark

Office and knowledge work

Claude Opus 5.5

GDPval-AA v2.1, which scores real professional deliverables: 1846 Elo against 1735 for Fable 5.1 and 1542 for GPT-6 Astra.

Vendor-reported benchmark

Business automation

GPT-6 Astra

AutomationBench: 41.4%, narrowly ahead of Opus 5.5 at 40.0%. One of the few results in Anthropic’s own table that Anthropic does not win.

Vendor-reported benchmark

Speed and price

Gemini 3.8 Flash

286.9 output tokens per second at $0.75 / $3.75 per million tokens — around five times faster than the top three and a fraction of their price, for a lower index score (41).

Independent measurement

Best model you can download

MiMo-V2.6-Pro

Highest-scoring open-weights model at 46, ahead of GLM-5.3 (45) and Kimi K3 (44). At about one trillion parameters it needs a server, not a desktop PC.

Independent measurement

Overall ranking

Artificial Analysis Intelligence Index v4.3.2 combines ten evaluations — coding, reasoning, science, knowledge work and tool use — which Artificial Analysis runs itself. Each model is shown at its highest-effort setting. Prices are US dollars per million input / output tokens.

#ModelReleasedIndexPriceSpeedWeights
1Claude Opus 5.5Anthropic22 Sep 202658$4 / $20Not yet measuredHosted only
2Claude Fable 5.1Anthropic1 Sep 202653$10 / $5065.7 tokens/sHosted only
3GPT-6 AstraOpenAI3 Sep 202653$10 / $5058.1 tokens/sHosted only
4MiMo-V2.6-ProXiaomi46Downloadable
5GLM-5.3Z AI45Downloadable
6Grok 4.6xAI12 Aug 202644$2 / $658.6 tokens/sHosted only
7Kimi K3Moonshot AI (Kimi)44Downloadable
8Gemini 3.8 FlashGoogle2 Sep 202641$0.75 / $3.75286.9 tokens/sHosted only

“—” means the figure is not published on the source page, not that it is zero. Scores from different index versions are not comparable, so this table uses only v4.3.2.

Head to head: Anthropic’s launch figures

From Anthropic’s Opus 5.5 announcement. These are vendor-reported: the company releasing the model chose the tests and ran them. Treat them as a claim to check against independent results, not as a verdict.

BenchmarkOpus 5.5Fable 5.1GPT-6 AstraOpus 5
Terminal-Bench 4.0Agentic coding in a terminal66.4%55.8%57.9%52.3%
FrontierCode v1.1Hard software tasks54.4%50.3%53.3%48.0%
GDPval-AA v2.1Professional deliverables (Elo)1846173515421708
AutomationBenchBusiness workflows40.0%31.4%41.4%26.9%
Humanity’s Last ExamExpert reasoning, with tools67.7%65.6%57.2%63.6%
OSWorld 2.0Operating a computer81.8%80.7%Not reported74.0%

What this means if you run AI on your own PC

None of the top three can be downloaded; they only run on their makers’ servers. The best open-weights models are close behind, but at 750 billion to 2.8 trillion parameters they need data-centre hardware. What runs on a home graphics card is smaller models from the same families — and whether a given one fits depends on your video memory, not on this ranking.

Sources

All figures retrieved 23 September 2026. Leaderboards change as models are added and re-tested.