Skip to content

Leaderboard

Long Context Reasoning

2 models scored. Prices are the cheapest listed across all serving providers. Cost/run and Pts/$ are priced for the chat token profile — short turns, light history, little caching — an assistant in a product.

This benchmark has not been calibrated to a token profile, so the figures use the site default. Treat the ordering as indicative rather than tuned to Long Context Reasoning. Profile: 1.5k in · 500 out · 20% cache hit · 8k context.

Scores are facts as publicly reported by labs (see report links below and on each model page) and aggregated via models.dev. Long Context Reasoning is maintained by its own project — Model Pulse is not affiliated with or endorsed by it.

2 models
#
🥇Fugusakana
74.7
1M
🥈Fugu Ultrasakana
73.3
$5$30$0.0223,4051.05M

Score reports: console.sakana.ai — full source URLs are linked on each model page.

Cheapest scorer: Fugu Ultra at $5 input /M.