Leaderboard
Artificial Analysis Coding Indexofficial site ↗
27 models scored · metric: index. Prices are the cheapest listed across all serving providers. Cost/run and Pts/$ are priced for the chat token profile — short turns, light history, little caching — an assistant in a product.
A benchmark is priced for the workload it exercises, not a fixed input:output blend: an agent loop and a single hard question bill very differently on the same rate card. Profile: 1.5k in · 500 out · 20% cache hit · 8k context.
Scores are facts as publicly reported by labs (see report links below and on each model page) and aggregated via models.dev. Artificial Analysis Coding Index is maintained by its own project — Model Pulse is not affiliated with or endorsed by it.
- 1.Step 3.5 Flash 2603122,046
- 2.Step 3.5 Flash116,713
- 3.Qwen3-Coder 30B-A3B Instruct93,663
- 4.GLM-4.5-Air58,405
- 5.Llama-3.3-70B-Instruct57,450
Score reports: openrouter.ai · openrouter.ai · openrouter.ai · artificialanalysis.ai · openrouter.ai · openrouter.ai — full source URLs are linked on each model page.
Cheapest scorer: Llama-3.3-70B-Instruct at $0.05 input /M.