1LLM1
LeaderboardBy taskCompareMethodology
Weekly rankings

Best LLM for each task

One model is rarely best at everything. Pick the category that matches your work.

Coding
Claude Fable 5.1

Claude Fable 5.1 has the highest Artificial Analysis Coding Index of the models we track — an independent test of writing and fixing real code.

Writing
—

Drafting, editing and tone control for everyday business writing.

Reasoning & Math
Claude Opus 5.5

Claude Opus 5.5 has the highest Artificial Analysis Intelligence Index of the models we track, which combines ten hard tests of maths, science, coding and reasoning.

Speed
Gemini 3.8 Flash

Gemini 3.8 Flash produces answers faster than any other model we track, based on Artificial Analysis's measured median output speed.

Cost-efficiency
GPT-6 Luna

GPT-6 Luna gives the most benchmark score per dollar: its Intelligence Index divided by its list price beats every other model we track.

Long context
—

Working reliably across very long documents.

Multimodal (vision)
—

Understanding images, charts, screenshots and documents.

Open-source
GLM-5.3

GLM-5.3 has the highest Intelligence Index among models whose weights you can download and run yourself.

Get the weekly LLM rankings

One email a week: what moved, what's new, and which model to actually use for your work. No hype, no benchmark jargon.

Subscribe free
1LLM1

The #1 LLM leaderboard. Independent rankings of large language models, updated weekly.

Leaderboards

  • Overall
  • Best for coding
  • Best value
  • Best open-source

Site

  • Compare models
  • Methodology
  • Admin

More from LLM1

  • LLM Research Directory
  • Weekly newsletter
© 2026 LLM1. All rankings are our own.Rankings last updated Oct 3, 2026
Edit with