Hourly diffs
Changelog
Every detected change to the model landscape: releases, price moves, context windows, capabilities and deprecations. Also available as RSS or JSON.
Change type
Time range
via openrouter
What changed
- Input price$0.15/M$0.3/M▲ 100.0%
- Output price$0.6/M$1.20/M▲ 100.0%
- Cache read$0.0030/M$0.0060/M▲ 100.0%
via edenai
What changed
- Input price$1.12/M$0.581/M▼ 48.2%
- Output price$3.37/M$1.74/M▼ 48.2%
via kilo
What changed
- Context203K1.05M5.2×
- Output limit182K131K−28%
via openrouter
What changed
- Cache read—$0.23/M
via openrouter
What changed
- Cache read—$0.05/M
via kilo
What changed
- Cache read—$0.23/M
via kilo
What changed
- Context256K203K−21%
- Output limit33K183K5.6×
via openrouter
What changed
- Output limit131K944K7.2×
- Input price$1.09/M$1.40/M▲ 28.2%
- Output price$3.43/M$4.40/M▲ 28.2%
- Cache read$0.203/M$0.26/M▲ 28.2%
via kilo
What changed
- Context1.05M1.02M−2%
- Output limit393K384K−2%
via openrouter
What changed
- Cache read—$0.03/M
via edenai
What changed
- Input price$0.352/M$0.176/M▼ 50.0%
- Output price$1.06/M$0.528/M▼ 50.0%
via openrouter
What changed
- Output limit8K16K2×
- Input price$0.228/M$0.12/M▼ 47.3%
- Output price$0.91/M$0.24/M▼ 73.6%
via openrouter
What changed
- Input price$0.04/M$0.06/M▲ 50.0%
- Output price$0.08/M$0.12/M▲ 50.0%
- Cache read$0.0080/M$0.012/M▲ 50.0%
via kilo
What changed
- Context131K41K−69%
- Output limit8K16K2×
via openrouter
What changed
- Output limit393K384K−2%
- Input price$0.578/M$1.05/M▲ 82.4%
- Output price$1.73/M$3.16/M▲ 82.4%
- Cache read$0.018/M$0.035/M▲ 91.1%
via kilo
What changed
- Cache read—$0.05/M
via openrouter
What changed
- Output limit33K183K5.6×
- Input price$0.625/M$0.6/M▼ 4.0%
- Output price$3.13/M$2.40/M▼ 23.2%
- Cache read$0.188/M$0.12/M▼ 36.0%
via kilo
What changed
- Cache read—$0.03/M
via openrouter
What changed
- Input price$0.049/M$0.09/M▲ 83.4%
- Output price$0.098/M$0.18/M▲ 83.4%
- Cache read$0.0098/M$0.018/M▲ 83.4%
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit182K131K−28%
- Input price$0.6/M$0.683/M▲ 13.9%
- Output price$2/M$2.15/M▲ 7.4%
- Cache read$0.15/M$0.127/M▼ 15.4%
via kilo
What changed
- Context1.05M1.05M−0%
- Output limit131K944K7.2×
via amazon-bedrock
What changed
- last updated2025-12-022025-12-01
via openrouter
What changed
- nameGemma 3 12BGemma 3 12B IT
- knowledge2024-08-312024-08
- release date2025-03-132025-03-12
- last updated2025-03-132025-03-12
via edenai
What changed
- Input price$0.3/M$0.2/M▼ 33.3%
- Output price$1.20/M$0.6/M▼ 50.0%
via merge-gateway
What changed
First observed in the model catalog
via ofox
What changed
- Output price$2/M$2.20/M▲ 10.0%
- Cache read$0.08/M$0.11/M▲ 37.5%
via ofox
What changed
- Input price$0.44/M$0.19/M▼ 56.8%
- Output price$1.32/M$0.51/M▼ 61.4%
- Cache read$0.014/M$0.028/M▲ 100.0%
via amazon-bedrock
What changed
- nameQwen3 Coder 30B A3B InstructQwen3-Coder 30B-A3B Instruct
- descriptionQwen coding model for software agents, repository edits, and code reasoningSmaller Qwen coder for efficient local agents and repo-level fixes
- open weightsNoYesEnabled
- knowledge2024-042025-04
- +1 more changes
via llmgateway-providers
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
- tool callNoYesEnabled
via tinfoil
What changed
- attachmentYesNoRemoved
via merge-gateway
What changed
- release date2024-11-012025-12-02
via amazon-bedrock
What changed
- descriptionCompact Mistral model for edge, latency-sensitive, and cost-efficient workloadsOpen vision-language model for efficient local deployment, instruction following, and tool use
- attachmentNoYesEnabled
- open weightsNoYesEnabled
- release date2024-12-012025-12-02
- +2 more changes
via llmgateway
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionOpen multimodal Llama model for long-context analysis and efficient agentsOpen Llama with long-context vision for efficient multimodal agents
- Context3.50M10M2.9×
- Output limit16K8K−50%
via agentrouter
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via deepinfra
What changed
First observed in the model catalog
via edenai
What changed
- Cache read$3/M—
- tiers[object Object][object Object]
via kilo
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- nameQwen/Qwen3-VL-235B-A22B-InstructQwen3 VL 235B A22B Instruct
- descriptionQwen vision-language model for visual reasoning, documents, and agent tasksQwen vision-language instruct model for visual reasoning, documents, and agent tasks
- open weightsNoYesEnabled
- knowledge—2025-03-31
- +4 more changes
via amazon-bedrock
What changed
- reasoningNoYesEnabled
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via llmgateway-providers
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionOpen-weight GPT model for self-hosted reasoning and instruction-following workloadsOpen GPT reasoning model for self-hosted agents and controllable deployments
via amazon-bedrock
What changed
- Context3.50M10M2.9×
- Output limit16K8K−50%
via llmgateway
What changed
First observed in the model catalog
via abacus
What changed
- attachmentYesNoRemoved
via kilo
What changed
- Context205K200K−2%
- Output limit131K128K−2%
via llmgateway-providers
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionCompact Mistral model for edge, latency-sensitive, and cost-efficient workloadsCompact open vision-language model for edge deployment, instruction following, and tool use
- attachmentNoYesEnabled
- open weightsNoYesEnabled
- release date2024-12-012025-12-02
- +2 more changes
via amazon-bedrock
What changed
- descriptionQwen coding model for software agents, repository edits, and code reasoningOpen-weight Qwen coding model for agents, repository edits, and multi-turn tool use
- reasoningYesNoRemoved
- knowledge—2025-09
- release date2026-02-062026-02-03
- +3 more changes
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via empiriolabs
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionFlagship Mistral model for advanced reasoning, coding, and multilingual workMistral's largest general model for enterprise agents, coding, and multilingual reasoning
- familymistralmistral-large
- attachmentNoYesEnabled
- knowledge—2024-11
via openrouter
What changed
- Context131K256K2×
via anyapi
What changed
- release date2024-11-012025-12-02
via kilo
What changed
- nameGoogle: Gemma 3 12BGemma 3 12B IT
- open weightsNoYesEnabled
- knowledge—2024-08
- release date2025-03-132025-03-12
- +1 more changes
via edenai
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.176/M$0.352/M▲ 100.0%
- Output price$0.528/M$1.06/M▲ 100.0%
via amazon-bedrock
What changed
- Input limit—1.04M
via openrouter
What changed
- Output limit944K236K−75%
- Input price$1/M$0.936/M▼ 6.4%
- Output price$3.41/M$3.17/M▼ 7.1%
- Cache read$0.2/M$0.187/M▼ 6.4%
via kilo
What changed
- nameGoogle: Gemma 3 27BGemma 3 27B IT
- open weightsNoYesEnabled
- knowledge—2024-08
via ofox
What changed
- Input price$0.4/M$0.6/M▲ 50.0%
- Output price$1.90/M$2.20/M▲ 15.8%
via amazon-bedrock
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- reasoningNoYesEnabled
via deepinfra
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionOpen multimodal Llama model for strong reasoning and fast responsesOpen multimodal Llama for strong reasoning with efficient everyday serving
- Output limit16K8K−50%
via amazon-bedrock
What changed
First observed in the model catalog
via llmgateway
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit384K393K1×
- Input price$1.05/M$0.578/M▼ 44.9%
- Output price$3.15/M$1.73/M▼ 44.9%
- Cache read$0.035/M$0.018/M▼ 47.4%
via amazon-bedrock
What changed
- structured outputNoYesEnabled
via venice
What changed
- knowledge—2024-08
via llmgateway-providers
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- familymistralpixtral
- attachmentNoYesEnabled
via amazon-bedrock
What changed
- descriptionOpen Llama instruction model for multilingual chat, reasoning, and codingPopular open Llama workhorse for multilingual chat, coding, and self-hosting
via llmgateway
What changed
First observed in the model catalog
via openrouter
What changed
- familymistralvoxtral
- release date2025-10-302025-07-15
- last updated2025-10-302025-07-15
via hyper
What changed
- Output price$2.78/M$2.75/M▼ 1.1%
via amazon-bedrock
What changed
- nameQwen3 Coder 480B A35B InstructQwen3-Coder 480B-A35B Instruct
- descriptionQwen coding model for software agents, repository edits, and code reasoningOpen Qwen coding heavyweight for repository reasoning and agentic engineering
- knowledge2024-042025-04
- release date2025-09-182025-07-23
- +1 more changes
via kilo
What changed
- nameDeepSeek V4 Flash LatestDeepSeek: DeepSeek V4 Flash Latest
- Output limit393K131K−67%
- Input price$0.05/M$0.035/M▼ 29.6%
- Output price$0.16/M$0.106/M▼ 34.0%
- +1 more changes
via ofox
What changed
- Input price$0.15/M$0.11/M▼ 26.7%
- Output price$0.47/M$0.39/M▼ 17.0%
- Cache read$0.016/M$0.011/M▼ 31.3%
- cache write$0.2/M$0.14/M▼ 30.0%
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit100K98K−2%
via kilo
What changed
- structured outputNoYesEnabled
- Input price$0.085/M$0.08/M▼ 5.9%
- Output price$0.4/M$0.45/M▲ 12.5%
via openrouter
What changed
- Output limit131K182K1.4×
- Input price$0.966/M$0.6/M▼ 37.9%
- Output price$3.04/M$2/M▼ 34.1%
- Cache read$0.193/M$0.15/M▼ 22.4%
via openrouter
What changed
First observed in the model catalog
via pioneer
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via merge-gateway
What changed
First observed in the model catalog
via kilo
What changed
Removed from the model catalog
via requesty
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via nan
What changed
Removed from the model catalog
via amazon-bedrock
What changed
- nameMiniMax M2.1MiniMax-M2.1
- descriptionMiniMax model for chat, coding, office work, and agentic tasksEarlier MiniMax agent model for practical coding and productivity tasks
- structured outputNoYesEnabled
via hyper
What changed
- Input price$0.404/M$0.396/M▼ 2.0%
- Output price$1.50/M$1.46/M▼ 2.1%
- Cache read$0.202/M$0.198/M▼ 2.0%
via amazon-bedrock
What changed
- attachmentNoYesEnabled
- open weightsNoYesEnabled
- release date2024-12-012025-10-28
- last updated2024-12-012025-10-28
- +1 more changes
via amazon-bedrock
What changed
- descriptionOpen Gemma instruction model for efficient chat and self-hosted deploymentsOpen multimodal Gemma instruction model for efficient text generation and image understanding
- attachmentNoYesEnabled
- tool callYesNoRemoved
- open weightsNoYesEnabled
- +4 more changes
via amazon-bedrock
What changed
- open weightsNoYesEnabled
- release date2024-12-012025-08-18
- last updated2024-12-012025-08-18
- Context128K131K1×
- +1 more changes
via openrouter
What changed
- Output limit100K98K−2%
via amazon-bedrock
What changed
- last updated2025-12-022025-12-01
via openrouter
What changed
- nameGoogle Gemini Flash LatestGemini Flash Latest
via amazon-bedrock
What changed
- descriptionMistral coding agent model for repository tasks and software engineering workflowsMistral's coding-agent model for repository work, terminal tasks, and software fixes
- knowledge—2025-12
- release date2026-02-172025-12-09
- last updated2026-02-172025-12-09
via nano-gpt
What changed
First observed in the model catalog
via ofox
What changed
- Input price$0.6/M$0.43/M▼ 28.3%
- Output price$3.60/M$2.57/M▼ 28.6%
via amazon-bedrock
What changed
- tool callYesNoRemoved
via edenai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Context1M1.05M1.1×
via openrouter
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionDeepSeek chat model for instruction following, coding, and analysisHybrid-reasoning DeepSeek model with thinking and non-thinking modes
- knowledge2024-07—
- release date2025-09-182025-08-21
via ofox
What changed
- Cache read$0.046/M$0.023/M▼ 50.0%
via vercel
What changed
- descriptionDeepSeek reasoning model for multi-step analysis, math, coding, and toolsClassic open reasoning model for transparent math, coding, and deliberate problem solving
- tool callYesNoRemoved
- open weightsNoYesEnabled
via kilo
What changed
- nameGoogle: Gemma 3 4BGemma 3 4B IT
- open weightsNoYesEnabled
- knowledge—2024-08
- release date2025-03-132025-03-12
- +1 more changes
via llmgateway
What changed
- release date2024-11-012025-12-02
via amazon-bedrock
What changed
- descriptionEfficient Mistral model for fast chat, extraction, and production assistantsOpen audio-language model for speech transcription, audio understanding, and voice-driven tool use
- familymistralvoxtral
- release date2025-07-012025-07-15
- last updated2025-07-012025-07-15
- +3 more changes
via cortecs
What changed
- namegemma-3-27b-itGemma 3 27B IT
- family—gemma
- open weightsNoYesEnabled
- knowledge—2024-08
via aihubmix
What changed
First observed in the model catalog
via requesty
What changed
- Input price$5.50/M$4.40/M▼ 20.0%
- Output price$33/M$22/M▼ 33.3%
- Cache read$0.55/M$0.44/M▼ 20.0%
via amazon-bedrock
What changed
- nameDeepSeek-V3.2DeepSeek V3.2
- descriptionDeepSeek chat model for instruction following, coding, and analysisHybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
- release date2026-02-062025-12-01
via pioneer
What changed
- attachmentYesNoRemoved
via amazon-bedrock
What changed
- descriptionFlagship GLM model for hybrid reasoning, coding, and agentic engineeringGeneral GLM flagship for coding, analysis, and tool-heavy engineering workflows
- release date2026-03-182026-02-12
- Output limit101K131K1.3×
via huggingface
What changed
First observed in the model catalog
via openrouter
What changed
- nameOpenAI GPT Mini LatestGPT Mini Latest
via llmgateway-providers
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionReasoning model for deliberate analysis, multi-step problem solving, and tool useEnterprise language model for workflow automation, coding, data analysis, and tool use
- release date2025-04-282024-10-09
via hyper
What changed
- Input price$0.274/M$0.255/M▼ 6.9%
- Output price$0.899/M$0.837/M▼ 7.0%
- Cache read$0.137/M$0.128/M▼ 6.9%
via kilo
What changed
First observed in the model catalog
via ofox
What changed
- Input price$0.05/M$0.043/M▼ 14.0%
via amazon-bedrock
What changed
First observed in the model catalog
via edenai
What changed
- release date2024-11-012025-12-02
via ofox
What changed
- Cache read$0.025/M$0.01/M▼ 60.0%
via ofox
What changed
- Output price$0.43/M$0.4/M▼ 7.0%
- Cache read$0.015/M$0.01/M▼ 33.3%
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.3/M$0.15/M▼ 50.0%
- Output price$1.20/M$0.6/M▼ 50.0%
- Cache read$0.0060/M$0.0030/M▼ 50.0%
via edenai
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- nameMiniMax M2MiniMax-M2
- descriptionMiniMax model for chat, coding, office work, and agentic tasksEfficient open MiniMax model built for coding agents and tool-heavy workflows
- structured outputNoYesEnabled
via openrouter
What changed
- Output limit33K236K7.2×
- Input price$0.042/M$0.09/M▲ 114.3%
- Output price$0.22/M$0.3/M▲ 36.4%
- Cache read—$0.05/M
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit384K393K1×
- Input price$0.955/M$1.60/M▲ 67.5%
- Output price$1.91/M$3.20/M▲ 67.5%
- Cache read$0.08/M$0.135/M▲ 69.6%
via amazon-bedrock
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via above
What changed
- descriptionOfficial DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decodingDeepSeek V4.1 Flash model for reasoning and agentic coding
- attachmentNoYesEnabled
- release date2026-07-312026-09-10
- last updated2026-07-312026-09-10
- +5 more changes
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.44/M$0.22/M▼ 50.0%
- Output price$1.32/M$0.66/M▼ 50.0%
- Cache read$0.014/M$0.0070/M▼ 50.0%
via kilo
What changed
- Output limit100K98K−2%
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- release date2024-11-012025-12-02
via kilo
What changed
- Output limit944K131K−86%
via edenai
What changed
- Input price$0.755/M$0.753/M▼ 0.2%
- Output price$0.755/M$0.753/M▼ 0.2%
via kilo
What changed
- nameMistral: Voxtral Small 24B 2507Voxtral Small 24B 2507
- familymistralvoxtral
- open weightsNoYesEnabled
- release date2025-10-302025-07-15
- +1 more changes
via fireworks-ai
What changed
- release date2024-11-012025-12-02
via amazon-bedrock
What changed
- nameQwen3 235B A22B 2507Qwen3 235B-A22B Instruct 2507
- descriptionQwen instruction model for multilingual chat, reasoning, and tool useUpdated large open Qwen3 MoE instruct model for multilingual chat, coding, and tool use
- knowledge2024-04—
- release date2025-09-182025-07-21
via amazon-bedrock
What changed
- last updated2025-12-022025-12-01
via amazon-bedrock
What changed
- Input limit—1.04M
via llmgateway-providers
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionCompact Mistral model for edge, latency-sensitive, and cost-efficient workloadsCompact open vision-language model for edge deployment, instruction following, and tool use
- attachmentNoYesEnabled
via crusoe
What changed
- attachmentYesNoRemoved
via bothub
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.465/M$0.464/M▼ 0.2%
- Output price$0.929/M$0.927/M▼ 0.2%
via volcengine-coding-plan
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via huggingface
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- nameGoogle Gemini Flash LatestGoogle: Gemini Flash Latest
via merge-gateway
What changed
First observed in the model catalog
via kilo
What changed
- nameAnthropic Claude Sonnet LatestAnthropic: Claude Sonnet Latest
via nano-gpt
What changed
- Input price$0.3/M$0.15/M▼ 50.0%
- Output price$1.20/M$0.6/M▼ 50.0%
- Cache read$0.0060/M$0.0030/M▼ 50.0%
via pioneer
What changed
First observed in the model catalog
via pioneer
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionFlagship GLM model for hybrid reasoning, coding, and agentic engineeringMature GLM model for dependable coding, reasoning, and structured agent tasks
via kilo
What changed
- Context1.02M1.05M1×
- Output limit384K393K1×
via kilo
What changed
- nameGoogle Gemini Pro LatestGoogle: Gemini Pro Latest
via llmgateway
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.08/M$0.065/M▼ 18.8%
- Output price$0.2/M$0.18/M▼ 10.0%
- Cache read$0.04/M—
via nano-gpt
What changed
First observed in the model catalog
via cortecs
What changed
- release date2024-11-012025-12-02
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- nameAnthropic Claude Haiku LatestAnthropic: Claude Haiku Latest
via amazon-bedrock
What changed
- nameQwen3 32B (dense)Qwen3 32B
- descriptionQwen instruction model for multilingual chat, reasoning, and tool useDense open Qwen model for self-hosted chat, reasoning, and coding
- knowledge2024-042025-04
- release date2025-09-182025-04
- +1 more changes
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Context1.02M1.05M1×
- Output limit384K393K1×
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via amazon-bedrock
What changed
- structured outputNoYesEnabled
via kilo
What changed
First observed in the model catalog
via nan
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- Output limit16K8K−50%
via amazon-bedrock
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionMistral reasoning model for transparent analysis, math, and complex decisionsOpen multimodal reasoning model for transparent analysis of text and images
- attachmentNoYesEnabled
- release date2025-12-022025-09-18
- last updated2025-12-022025-09-18
via ofox
What changed
- Input price$2/M$1.71/M▼ 14.5%
- Output price$6/M$5.14/M▼ 14.3%
- Cache read$0.25/M$0.17/M▼ 32.0%
- cache write$2.50/M$2.14/M▼ 14.4%
via merge-gateway
What changed
First observed in the model catalog
via kilo
What changed
- Context1M203K−80%
- Output limit131K182K1.4×
via requesty
What changed
First observed in the model catalog
via openrouter
What changed
- namegpt-oss-safeguard-20bGPT OSS Safeguard 20B
via cortecs
What changed
- Input price$0.13/M$0.09/M▼ 30.8%
- Output price$0.28/M$0.17/M▼ 39.3%
- Cache read$0.03/M$0.014/M▼ 53.3%
via openrouter
What changed
- Output limit236K32K−86%
- Input price$0.09/M$0.048/M▼ 46.5%
- Output price$0.3/M$0.193/M▼ 35.6%
via openrouter
What changed
Removed from the model catalog
via deepinfra
What changed
First observed in the model catalog
via llmgateway
What changed
- Input price$0.14/M$0.132/M▼ 5.7%
- Output price$0.58/M$0.528/M▼ 9.0%
- Cache read$0.035/M$0.033/M▼ 5.7%
via kilo
What changed
First observed in the model catalog
via togetherai
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionReasoning model for deliberate analysis, multi-step problem solving, and tool useEnterprise language model for workflow automation, coding, data analysis, and tool use
- release date2025-04-282024-10-09
via amazon-bedrock
What changed
- descriptionKimi multimodal agent model for visual understanding, coding, and planningEarlier Kimi frontier model for long-context agents, coding, and multimodal work
- familykimikimi-k2
- attachmentNoYesEnabled
- knowledge—2025-01
- +2 more changes
via kilo
What changed
- nameMoonshotAI Kimi LatestMoonshotAI: Kimi Latest
- Input price$2.34/M$2.10/M▼ 10.3%
- Output price$11.70/M$10.95/M▼ 6.4%
- Cache read$0.261/M—
via llmgateway-providers
What changed
First observed in the model catalog
via huggingface
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionOpen-weight GPT model for self-hosted reasoning and instruction-following workloadsOpen GPT reasoning model for self-hosted agents and controllable deployments
via kilo
What changed
- Context1.05M262K−75%
- Output limit944K236K−75%
- Input price$1/M$0.936/M▼ 6.4%
- Output price$3.41/M$3.17/M▼ 7.1%
- +1 more changes
via kilo
What changed
First observed in the model catalog
via greenpt
What changed
First observed in the model catalog
via ofox
What changed
- Input price$2/M$1.71/M▼ 14.5%
- Output price$6/M$5.14/M▼ 14.3%
- Cache read$0.25/M$0.17/M▼ 32.0%
- cache write$2.50/M$2.14/M▼ 14.4%
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$5.25/M$5/M▼ 4.8%
- Output price$31.50/M$30/M▼ 4.8%
- Cache read$0.525/M$0.5/M▼ 4.8%
via amazon-bedrock
What changed
- release date2025-12-232025-12-15
- Context128K262K2×
- Output limit4K8K2×
via openrouter
What changed
- nameGemma 3 4BGemma 3 4B IT
- knowledge2024-08-312024-08
- release date2025-03-132025-03-12
- last updated2025-03-132025-03-12
via empiriolabs
What changed
First observed in the model catalog
via kilo
What changed
First observed in the model catalog
via cortecs
What changed
- Input price$0.334/M$0.1/M▼ 70.1%
- Output price$2.45/M$0.4/M▼ 83.7%
- Cache read$0.111/M$0.04/M▼ 64.0%
via openrouter
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.2%
- Output price$0.755/M$0.753/M▼ 0.2%
via amazon-bedrock
What changed
- descriptionOpen Llama instruction model for multilingual chat, reasoning, and codingCompact open Llama model for lightweight chat, drafting, and self-hosting
via openrouter
What changed
- Output limit33K472K14.4×
via hyper
What changed
- Input price$0.178/M$0.168/M▼ 5.6%
- Output price$0.68/M$0.66/M▼ 2.9%
- Cache read$0.089/M$0.084/M▼ 5.6%
via openrouter
What changed
- structured outputNoYesEnabled
via kilo
What changed
Removed from the model catalog
via amazon-bedrock
What changed
- descriptionEfficient Mistral model for fast chat, extraction, and production assistantsOpen audio-language model for speech transcription, audio understanding, and voice-driven tool use
- familymistralvoxtral
- attachmentNoYesEnabled
- open weightsNoYesEnabled
- +4 more changes
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.2%
- Output price$0.697/M$0.696/M▼ 0.2%
via nano-gpt
What changed
- Input price$5.25/M$5/M▼ 4.8%
- Output price$31.50/M$30/M▼ 4.8%
- Cache read$0.525/M$0.5/M▼ 4.8%
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit393K131K−67%
- Input price$0.05/M$0.035/M▼ 29.6%
- Output price$0.16/M$0.106/M▼ 34.0%
- Cache read$0.013/M$0.0011/M▼ 91.4%
via amazon-bedrock
What changed
- nameMiniMax M2.5MiniMax-M2.5
- descriptionMiniMax model for chat, coding, office work, and agentic tasksPrior MiniMax coding model for agent workflows, office edits, and automation
- structured outputNoYesEnabled
- release date2026-03-182026-02-12
via amazon-bedrock
What changed
- descriptionOpen-weight GPT model for self-hosted reasoning and instruction-following workloadsOpen GPT reasoning model for self-hosted agents and controllable deployments
via pioneer
What changed
First observed in the model catalog
via llmgateway
What changed
First observed in the model catalog
via kilo
What changed
- Context262K128K−51%
- Output limit236K32K−86%
via openrouter
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via kilo
What changed
- Context1M262K−74%
via kilo
What changed
- Output limit100K98K−2%
via digitalocean
What changed
- Output limit1.05M128K−88%
- Input price$1.40/M$0.95/M▼ 32.1%
- Output price$4.40/M$3.40/M▼ 22.7%
- Cache read$0.26/M$0.2/M▼ 23.1%
via kilo
What changed
- Context1.05M524K−50%
- Output limit33K472K14.4×
via ofox
What changed
- Input price$2.50/M$1.71/M▼ 31.6%
- Output price$7.50/M$5.14/M▼ 31.5%
- Cache read$0.5/M$0.17/M▼ 66.0%
- cache write$3.13/M$2.14/M▼ 31.5%
via edenai
What changed
- Input price$0.581/M$1.12/M▲ 93.2%
- Output price$1.74/M$3.37/M▲ 93.2%
via edenai
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- last updated2025-12-022025-12-01
via openrouter
What changed
- structured outputYesNoRemoved
via snowflake-cortex
What changed
- attachmentYesNoRemoved
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionOpen-weight GPT model for self-hosted reasoning and instruction-following workloadsOpen GPT reasoning model for self-hosted agents and controllable deployments
via watsonx
What changed
- attachmentYesNoRemoved
via cline-pass
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.122/M$0.102/M▼ 16.4%
- Output price$0.42/M$0.356/M▼ 15.2%
- Cache read$0.061/M$0.051/M▼ 16.4%
via edenai
What changed
First observed in the model catalog
via evroc
What changed
- attachmentYesNoRemoved
via openrouter
What changed
- Input price$0.065/M$0.04/M▼ 38.5%
- Output price$0.18/M$0.08/M▼ 55.6%
- Cache read$0.016/M$0.0080/M▼ 50.0%
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
- release date2026-02-232025-10-29
via kilo
What changed
- Output limit8K16K2×
via edenai
What changed
First observed in the model catalog
via opper
What changed
- release date2024-11-012025-12-02
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- Context131K262K2×
- Output limit33K236K7.2×
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
- nameMoonshotAI Kimi LatestKimi Latest
- Input price$2.34/M$2.10/M▼ 10.3%
- Output price$11.70/M$10.95/M▼ 6.4%
- Cache read$0.261/M—
via openrouter
What changed
- Output limit8K16K2×
- Input price$0.72/M$0.4/M▼ 44.4%
- Output price$0.72/M$0.4/M▼ 44.4%
via kilo
What changed
First observed in the model catalog
via baseten
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit944K131K−86%
- Input price$1.40/M$1.09/M▼ 22.0%
- Output price$4.40/M$3.43/M▼ 22.0%
- Cache read$0.26/M$0.203/M▼ 22.0%
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.3/M$0.5/M▲ 66.7%
- Output price$1.20/M$1.50/M▲ 25.0%
- Cache read$0.0060/M$0.05/M▲ 733.3%
via edenai
What changed
- Cache read$3/M—
- tiers[object Object][object Object]
via cortecs
What changed
- Input price$0.201/M$0.1/M▼ 50.2%
- Output price$0.5/M$0.35/M▼ 30.0%
- Cache read$0.05/M$0.018/M▼ 64.0%
via openrouter
What changed
- Input price$0.42/M$0.214/M▼ 49.0%
- Output price$3/M$2.55/M▼ 15.0%
- Cache read$0.085/M$0.15/M▲ 76.5%
via openrouter
What changed
- Input price$0.29/M$0.25/M▼ 13.8%
- Output price$1.14/M$1/M▼ 12.3%
- Cache read$0.11/M—
via llmgateway-providers
What changed
- release date2024-11-012025-12-02
via merge-gateway
What changed
- Output limit250K128K−49%
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via merge-gateway
What changed
- Context1M1.05M1×
- Input price$0.22/M$0.035/M▼ 84.1%
- Output price$0.66/M$0.07/M▼ 89.4%
via amazon-bedrock
What changed
- nameGoogle Gemma 3 12BGemma 3 12B IT
- descriptionOpen Gemma instruction model for efficient chat and self-hosted deploymentsOpen multimodal Gemma instruction model for multilingual text generation and image understanding
- attachmentNoYesEnabled
- open weightsNoYesEnabled
- +5 more changes
via amazon-bedrock
What changed
- nameQwen/Qwen3-Next-80B-A3B-InstructQwen3-Next 80B-A3B Instruct
- open weightsNoYesEnabled
- knowledge—2025-04
- release date2025-09-182025-09-11
- +3 more changes
via kilo
What changed
- Output limit16K236K14.4×
via llmgateway
What changed
- attachmentYesNoRemoved
via ofox
What changed
- Input price$0.45/M$0.5/M▲ 11.1%
- Output price$3.20/M$1.71/M▼ 46.6%
- Cache read$0.05/M$0.043/M▼ 14.0%
- cache write$0.563/M$0.63/M▲ 12.0%
via edenai
What changed
- Cache read$3/M—
- tiers[object Object][object Object]
via ofox
What changed
- Cache read$0.044/M$0.15/M▲ 240.9%
via kilo
What changed
- Input price$0.29/M$0.25/M▼ 13.8%
- Output price$1.14/M$1/M▼ 12.3%
- Cache read$0.11/M—
via openrouter
What changed
- release date2024-11-012025-12-02
via amazon-bedrock
What changed
- last updated2025-12-022025-12-01
via kilo
What changed
- structured outputYesNoRemoved
via kilo
What changed
- Output limit100K98K−2%
via agentrouter
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- descriptionDeepSeek reasoning model for multi-step analysis, math, coding, and toolsClassic open reasoning model for transparent math, coding, and deliberate problem solving
- tool callYesNoRemoved
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.3/M$0.15/M▼ 50.0%
- Output price$1.20/M$0.6/M▼ 50.0%
- Cache read$0.0060/M$0.0030/M▼ 50.0%
via openrouter
What changed
- Output limit131K128K−2%
- Input price$0.3/M$0.27/M▼ 10.0%
- Output price$1.20/M$1.08/M▼ 10.0%
- Cache read$0.03/M$0.027/M▼ 10.0%
via amazon-bedrock
What changed
- descriptionKimi reasoning model for long-horizon research, planning, and tool useThinking Kimi model for slower research passes, planning, and hard technical questions
- knowledge—2024-08
- release date2025-12-022025-11-06
via cortecs
What changed
- last updated2025-12-022025-12-01
via ollama-cloud
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.3/M$0.35/M▲ 16.7%
- Output price$1.10/M$1.50/M▲ 36.4%
via nvidia
What changed
First observed in the model catalog
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
- nameAnthropic Claude Haiku LatestClaude Haiku Latest
via amazon-bedrock
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via hyper
What changed
- Input price$1.33/M$1.32/M▼ 0.6%
- Output price$4.22/M$4.31/M▲ 2.1%
- Cache read$0.663/M$0.659/M▼ 0.6%
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via pioneer
What changed
- release date2024-11-012025-12-02
via amazon-bedrock
What changed
- nameGoogle Gemma 3 27B InstructGemma 3 27B IT
- descriptionOpen Gemma instruction model for efficient chat and self-hosted deploymentsLargest open Gemma 3 instruction model for multilingual text generation and visual understanding
- tool callYesNoRemoved
- knowledge2025-072024-08
- +4 more changes
via amazon-bedrock
What changed
- descriptionEfficient GLM model for fast reasoning, coding, and agent workflowsBudget GLM lane for fast coding help, routing, and everyday automation
via openrouter
What changed
- Output limit16K236K14.4×
- Input price$0.22/M$0.087/M▼ 60.2%
- Output price$0.88/M$0.35/M▼ 60.2%
- Cache read—$0.018/M
via llmgateway-providers
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- nameGoogle Gemini Pro LatestGemini Pro Latest
via openrouter
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.089/M$0.049/M▼ 44.7%
- Output price$0.177/M$0.098/M▼ 44.7%
- Cache read$0.018/M$0.0098/M▼ 44.7%
via requesty
What changed
Removed from the model catalog
via openrouter
What changed
- nameGemma 3 27BGemma 3 27B IT
- knowledge2024-08-312024-08
via openrouter
What changed
- nameAnthropic Claude Sonnet LatestClaude Sonnet Latest
via openrouter
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit100K98K−2%
via llmgateway-providers
What changed
First observed in the model catalog
via pioneer
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via ollama-cloud
What changed
First observed in the model catalog
via venice
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via kilo
What changed
- nameOpenAI GPT Mini LatestOpenAI: GPT Mini Latest
via llmgateway-providers
What changed
First observed in the model catalog
via nan
What changed
- nameDeepSeek V4 FlashDeepSeek V4.1 Flash
- descriptionFast DeepSeek V4 lane for economical reasoning, coding, and long-context workDeepSeek V4.1 Flash model for reasoning and agentic coding
- attachmentNoYesEnabled
- release date2026-04-242026-09-10
- +2 more changes
via pioneer
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$2.34/M$2.65/M▲ 13.2%
- Output price$11.70/M$13.28/M▲ 13.5%
- Cache read$0.261/M$0.303/M▲ 16.0%
via pioneer
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via volcengine-coding-plan
What changed
First observed in the model catalog
via pioneer
What changed
First observed in the model catalog
via edenai
What changed
- Input price$1.05/M$1.04/M▼ 0.2%
- Output price$1.05/M$1.04/M▼ 0.2%
via kilo
What changed
- nameOpenAI: gpt-oss-safeguard-20bGPT OSS Safeguard 20B
- open weightsNoYesEnabled
via baseten
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via ofox
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via google-vertex
What changed
First observed in the model catalog
via nano-gpt
What changed
- descriptionDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.DeepSeek V4.1 Flash model for reasoning and agentic coding
- familydeepseekdeepseek-flash
- open weightsNoYesEnabled
- knowledge—2025-05
- +5 more changes
via amazon-bedrock
What changed
- Output limit8K10K1.2×
- cache write—$0.04/M
via kilo
What changed
- Output limit16K8K−50%
via ofox
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.07/M$0.075/M▲ 7.1%
- Output price$0.233/M$0.25/M▲ 7.2%
- Cache read$0.014/M$0.015/M▲ 7.1%
via ofox
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via scnet-token-plan
What changed
Removed from the model catalog
via scnet-token-plan
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.175/M$0.174/M▼ 0.3%
- Output price$0.757/M$0.755/M▼ 0.3%
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.92/M
via hyper
What changed
- Input price$0.18/M$0.178/M▼ 1.1%
- Output price$0.61/M$0.68/M▲ 11.5%
- Cache read—$0.089/M
- cache write$0.09/M—
via sap-ai-core
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.064/M
via openrouter
What changed
- Input price$0.07/M$0.15/M▲ 114.3%
- Output price$0.233/M$0.5/M▲ 114.3%
- Cache read$0.014/M$0.03/M▲ 114.3%
via ofox
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.069/M
via vercel
What changed
First observed in the model catalog
via vercel
What changed
Removed from the model catalog
via ofox
What changed
First observed in the model catalog
via crossmodel
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via bothub
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.8/M
via bothub
What changed
First observed in the model catalog
via edenai
What changed
- Output limit8K10K1.2×
via hyper
What changed
- Cache read—$0.059/M
- cache write$0.059/M—
via nano-gpt
What changed
- Input price$0.14/M$0.05/M▼ 64.3%
- Output price$0.28/M$0.16/M▼ 42.9%
- Cache read$0.014/M$0.013/M▼ 7.1%
via llmgateway-providers
What changed
First observed in the model catalog
via llmgateway
What changed
First observed in the model catalog
via llmgateway
What changed
Removed from the model catalog
via opencode-go
What changed
First observed in the model catalog
via hyper
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via llmgateway-providers
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- Output limit8K10K1.2×
- cache write—$0.037/M
via openrouter
What changed
- Output limit16K8K−50%
- Input price$0.4/M$0.72/M▲ 80.0%
- Output price$0.4/M$0.72/M▲ 80.0%
via llmgateway-providers
What changed
- Context524K1M1.9×
via google-vertex
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via huggingface
What changed
First observed in the model catalog
via hyper
What changed
- Cache read—$0.3/M
- cache write$0.3/M—
via amazon-bedrock
What changed
- knowledge—2025-10
- release date2025-12-012025-12-02
- last updated2025-12-012025-12-02
- Output limit64K66K1×
- +1 more changes
via fireworks-ai
What changed
First observed in the model catalog
via opencode-go
What changed
- Input price$0.22/M$0.15/M▼ 31.8%
- Output price$0.66/M$0.6/M▼ 9.1%
- Cache read$0.0070/M$0.0030/M▼ 57.1%
via ofox
What changed
First observed in the model catalog
via google-vertex
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via inception
What changed
First observed in the model catalog
via deepinfra
What changed
- status—deprecated
via deepinfra
What changed
- Input price$0.4/M$0.14/M▼ 65.0%
- Output price$2/M$0.28/M▼ 86.0%
- Cache read$0.08/M$0.0028/M▼ 96.5%
via amazon-bedrock
What changed
- Output limit8K10K1.2×
- cache write—$0.035/M
via vercel
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.8/M
via nano-gpt
What changed
- Input price$0.14/M$0.05/M▼ 64.3%
- Output price$0.28/M$0.16/M▼ 42.9%
- Cache read$0.014/M$0.013/M▼ 7.1%
via ofox
What changed
First observed in the model catalog
via edenai
What changed
- Output limit8K10K1.2×
via deepinfra
What changed
- status—deprecated
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.063/M
via ofox
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via ofox
What changed
- Input price$1.26/M$1.40/M▲ 11.1%
- Output price$3.96/M$4.40/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via ofox
What changed
First observed in the model catalog
via openrouter
What changed
- Context203K200K−1%
via hyper
What changed
- Input price$1.39/M$1.33/M▼ 4.3%
- Output price$4.36/M$4.22/M▼ 3.1%
- Cache read—$0.663/M
- cache write$0.693/M—
via deepseek
What changed
- descriptionOfficial DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decodingDeepSeek V4.1 Flash model for reasoning and agentic coding
- attachmentNoYesEnabled
- release date2026-07-312026-09-10
- last updated2026-07-312026-09-10
- +5 more changes
via ofox
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via llmgateway
What changed
- Input price$0.102/M$0.1/M▼ 2.0%
- Output price$0.297/M$0.25/M▼ 15.8%
- Cache read$0.012/M$0.01/M▼ 16.7%
via ofox
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.2/M$0.04/M▼ 80.0%
- Output price$0.8/M$0.15/M▼ 81.3%
- Cache read$0.1/M$0.02/M▼ 80.0%
via nano-gpt
What changed
- descriptionDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.DeepSeek V4.1 Flash model for reasoning and agentic coding
- familydeepseekdeepseek-flash
- open weightsNoYesEnabled
- knowledge—2025-05
- +5 more changes
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$2.40/M$2.34/M▼ 2.5%
- Output price$12/M$11.70/M▼ 2.5%
- Cache read$0.24/M$0.261/M▲ 8.8%
via ofox
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.16/M$0.14/M▼ 12.5%
- Output price$0.47/M$0.42/M▼ 10.6%
via hyper
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via hyper
What changed
- Cache read—$0.137/M
- cache write$0.137/M—
via vercel
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.06/M
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.8/M
via vercel
What changed
- Input price$0.28/M$0.62/M▲ 121.4%
- Output price$0.42/M$1.85/M▲ 340.5%
- Cache read$0.028/M—
via ofox
What changed
- Input price$0.98/M$1.40/M▲ 42.9%
- Output price$3.08/M$4.40/M▲ 42.9%
- Cache read$0.182/M$0.26/M▲ 42.9%
via ofox
What changed
First observed in the model catalog
via deepseek
What changed
- descriptionExperimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic workDeepSeek V4.1 Flash model for reasoning and agentic coding
- statusbeta—
- open weightsNoYesEnabled
- knowledge—2025-05
- +6 more changes
via edenai
What changed
- Input price$1.05/M$1.05/M▼ 0.3%
- Output price$1.05/M$1.05/M▼ 0.3%
via venice
What changed
First observed in the model catalog
via hyper
What changed
- Cache read—$0.223/M
- cache write$0.223/M—
via bothub
What changed
First observed in the model catalog
via google-vertex
What changed
First observed in the model catalog
via kilo
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.757/M$0.755/M▼ 0.3%
- Output price$0.757/M$0.755/M▼ 0.3%
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.84/M
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit16K33K2×
- Input price$0.07/M$0.042/M▼ 40.0%
- Output price$0.34/M$0.22/M▼ 35.3%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.14/M$0.05/M▼ 64.3%
- Output price$0.28/M$0.16/M▼ 42.9%
- Cache read$0.014/M$0.013/M▼ 7.1%
via alibaba-token-plan-cn
What changed
- statusbetadeprecated
via vercel
What changed
- Output limit8K10K1.2×
- cache write—$0.035/M
via openrouter
What changed
- Input price$0.03/M$0.09/M▲ 200.0%
- Output price$0.12/M$0.36/M▲ 200.0%
- Cache read$0.0060/M$0.018/M▲ 200.0%
via openrouter
What changed
- Input price$0.07/M$0.075/M▲ 7.1%
- Output price$0.233/M$0.25/M▲ 7.2%
- Cache read$0.014/M$0.015/M▲ 7.1%
via ofox
What changed
First observed in the model catalog
via cortecs
What changed
- knowledge—2025-10
- release date2025-12-012025-12-02
- last updated2025-12-012025-12-02
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.06/M
via llmgateway-providers
What changed
First observed in the model catalog
via edenai
What changed
- Output limit8K10K1.2×
via deepseek
What changed
First observed in the model catalog
via opencode-go
What changed
- status—deprecated
via ofox
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via google-vertex
What changed
First observed in the model catalog
via ofox
What changed
- Input price$0.075/M$0.15/M▲ 100.0%
- Output price$0.25/M$0.5/M▲ 100.0%
- Cache read$0.015/M$0.03/M▲ 100.0%
via openrouter
What changed
First observed in the model catalog
via edenai
What changed
- Output limit8K10K1.2×
via amazon-bedrock
What changed
- modalities.inputtext,image,videotext,image,video,pdf
- Output limit8K10K1.2×
- cache write—$0.06/M
via ofox
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.466/M$0.465/M▼ 0.3%
- Output price$0.932/M$0.929/M▼ 0.3%
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- Cache read$0.033/M$0.035/M▲ 6.1%
via amazon-bedrock
What changed
- knowledge—2025-10
- release date2025-12-012025-12-02
- last updated2025-12-012025-12-02
- Output limit64K66K1×
- +1 more changes
via amazon-bedrock
What changed
- attachmentNoYesEnabled
- knowledge—2025-10
- release date2024-12-012025-12-02
- last updated2024-12-012025-12-02
- +5 more changes
via ofox
What changed
First observed in the model catalog
via merge-gateway
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.175/M$0.174/M▼ 0.3%
- Output price$0.699/M$0.697/M▼ 0.3%
via bothub
What changed
First observed in the model catalog
via above
What changed
Removed from the model catalog
via alibaba-token-plan
What changed
- statusbetadeprecated
via kilo
What changed
- Context262K131K−50%
- Output limit16K33K2×
via amazon-bedrock
What changed
- knowledge—2025-10
- release date2025-12-012025-12-02
- last updated2025-12-012025-12-02
- Output limit64K66K1×
- +1 more changes
via hyper
What changed
- Input price$0.106/M$0.122/M▲ 15.1%
- Output price$0.368/M$0.42/M▲ 14.1%
- Cache read—$0.061/M
- cache write$0.053/M—
via edenai
What changed
- Output limit8K10K1.2×
via hyper
What changed
- Cache read—$0.279/M
- cache write$0.279/M—
via fireworks-ai
What changed
First observed in the model catalog
via pioneer
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$3/M$2.34/M▼ 22.0%
- Output price$15/M$11.70/M▼ 22.0%
- Cache read$0.3/M$0.261/M▼ 13.0%
via kilo
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- knowledge—2025-10
- release date2025-12-012025-12-02
- last updated2025-12-012025-12-02
- Output limit64K66K1×
- +1 more changes
via ofox
What changed
First observed in the model catalog
via kilo
What changed
- Input price$2.40/M$2.34/M▼ 2.5%
- Output price$12/M$11.70/M▼ 2.5%
- Cache read$0.24/M$0.261/M▲ 8.8%
via llmgateway-providers
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via opencode-go
What changed
- Input price$0.22/M$0.15/M▼ 31.8%
- Output price$0.66/M$0.6/M▼ 9.1%
- Cache read$0.0070/M$0.0030/M▼ 57.1%
via crossmodel
What changed
- Input price$0.405/M$0.27/M▼ 33.3%
- Output price$1.22/M$1.08/M▼ 11.1%
- Cache read$0.013/M$0.0054/M▼ 60.0%
- cache write$0.405/M$0.27/M▼ 33.3%
via ofox
What changed
First observed in the model catalog
via deepinfra
What changed
First observed in the model catalog
via crossmodel
What changed
- Input price$0.405/M$0.27/M▼ 33.3%
- Output price$1.22/M$1.08/M▼ 11.1%
- Cache read$0.013/M$0.0054/M▼ 60.0%
- cache write$0.405/M$0.27/M▼ 33.3%
via edenai
What changed
- Output limit8K10K1.2×
via deepinfra
What changed
- status—deprecated
via bothub
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- Output limit8K10K1.2×
- cache write—$0.035/M
via merge-gateway
What changed
- Context1M1.05M1×
- Input price$0.22/M$0.035/M▼ 84.1%
- Output price$0.66/M$0.07/M▼ 89.4%
via hyper
What changed
- Cache read—$0.43/M
- cache write$0.43/M—
via kilo
What changed
- Context1.02M1.05M1×
- Output limit128K944K7.4×
- Input price$1.11/M$1/M▼ 10.2%
- Output price$3.50/M$3.41/M▼ 2.5%
- +1 more changes
via openrouter
What changed
First observed in the model catalog
via kilo
What changed
- Context1.05M1M−5%
via venice
What changed
- Input price$0.25/M$0.05/M▼ 80.0%
- Output price$0.938/M$0.187/M▼ 80.0%
- Cache read$0.025/M$0.0050/M▼ 80.0%
via merge-gateway
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.458/M$0.404/M▼ 11.8%
- Output price$1.71/M$1.50/M▼ 12.6%
- Cache read—$0.202/M
- cache write$0.229/M—
via google-vertex
What changed
First observed in the model catalog
via hyper
What changed
- Cache read—$0.303/M
- cache write$0.303/M—
via ofox
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit128K944K7.4×
- Input price$1.11/M$1/M▼ 10.2%
- Output price$3.50/M$3.41/M▼ 2.5%
- Cache read$0.207/M$0.2/M▼ 3.2%
via llmgateway
What changed
- Input price$0.89/M$0.95/M▲ 6.7%
- Output price$3.71/M$4/M▲ 7.8%
- Cache read$0.18/M$0.19/M▲ 5.6%
via edenai
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via zenmux
What changed
- status—deprecated
via nearai
What changed
- Input price$1.80/M$1.75/M▼ 2.8%
- Output price$15.50/M$14/M▼ 9.7%
- Cache read$0.18/M$0.175/M▼ 2.8%
via venice
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via openrouter
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via openrouter
What changed
- Output limit944K131K−86%
- Input price$0.075/M$0.07/M▼ 6.7%
- Output price$0.25/M$0.233/M▼ 6.7%
- Cache read$0.015/M$0.014/M▼ 6.7%
via kilo
What changed
- Output limit236K100K−57%
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit472K33K−93%
via amazon-bedrock
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$1.99/M$2.20/M▲ 10.6%
- Output price$6.16/M$6.50/M▲ 5.5%
- Cache read$0.4/M$0.45/M▲ 12.5%
via amazon-bedrock
What changed
First observed in the model catalog
via opencode-go
What changed
- nameGLM-5.3-Flash (2x usage)GLM-5.3-Flash
- Input price$0.075/M$0.15/M▲ 100.0%
- Output price$0.25/M$0.5/M▲ 100.0%
- Cache read$0.015/M$0.03/M▲ 100.0%
via amazon-bedrock
What changed
First observed in the model catalog
via kilo
What changed
- Context1.05M1.02M−2%
- Output limit393K384K−2%
via edenai
What changed
First observed in the model catalog
via deepinfra
What changed
- Cache read$0.14/M$0.014/M▼ 90.0%
via nano-gpt
What changed
Removed from the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via nearai
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via nearai
What changed
- Output price$15.50/M$15/M▼ 3.2%
via merge-gateway
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$1.81/M$2/M▲ 10.2%
- Output price$5.45/M$6/M▲ 10.2%
- Cache read$0.21/M$0.25/M▲ 19.0%
via cortecs
What changed
- namenova-2-liteNova 2 Lite
- family—nova
- release date2025-12-042025-12-01
- last updated2025-12-042025-12-01
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$2.83/M$3.50/M▲ 23.7%
- Output price$14.13/M$18/M▲ 27.4%
- Cache read$0.28/M$0.35/M▲ 25.0%
via nearai
What changed
- Context256K16K−94%
- Output limit33K8K−75%
via edenai
What changed
- Input price$0.174/M$0.175/M▲ 0.3%
- Output price$0.697/M$0.699/M▲ 0.3%
via openrouter
What changed
- Input price$0.09/M$0.22/M▲ 144.4%
- Output price$0.55/M$0.88/M▲ 60.0%
via kilo
What changed
- Output limit82K66K−20%
via kilo
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via amazon-bedrock
What changed
- structured outputNoYesEnabled
via amazon-bedrock
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$0.89/M$0.95/M▲ 6.7%
- Output price$3.71/M$4/M▲ 7.8%
- Cache read$0.18/M$0.19/M▲ 5.6%
via wandb
What changed
- knowledge—2023-12
via openrouter
What changed
- Output limit16K16K−2%
- Input price$0.32/M$0.257/M▼ 19.6%
- Output price$0.89/M$1.03/M▲ 15.6%
via amazon-bedrock
What changed
First observed in the model catalog
via kilo
What changed
- Context524K1.05M2×
- Output limit472K33K−93%
via kilo
What changed
- Output limit944K131K−86%
- Input price$0.075/M$0.07/M▼ 6.7%
- Output price$0.25/M$0.233/M▼ 6.7%
- Cache read$0.015/M$0.014/M▼ 6.7%
via openrouter
What changed
- Output limit236K100K−57%
- Cache read—$0.15/M
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit82K66K−20%
- Input price$0.29/M$0.26/M▼ 10.3%
- Output price$2.40/M$2.08/M▼ 13.3%
via llmgateway
What changed
- Input price$1.81/M$2/M▲ 10.2%
- Output price$5.45/M$6/M▲ 10.2%
- Cache read$0.21/M$0.25/M▲ 19.0%
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- Context128K262K2×
- Output limit32K236K7.4×
via amazon-bedrock
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.075/M$0.07/M▼ 6.7%
- Output price$0.25/M$0.233/M▼ 6.7%
- Cache read$0.015/M$0.014/M▼ 6.7%
via llmgateway
What changed
- Input price$2.83/M$3/M▲ 6.0%
- Output price$14.13/M$15/M▲ 6.2%
- Cache read$0.28/M$0.3/M▲ 7.1%
via llmgateway-providers
What changed
- Input price$0.55/M$0.8/M▲ 45.5%
- Output price$1.78/M$2.55/M▲ 42.9%
- Cache read$0.111/M$0.16/M▲ 44.1%
via openrouter
What changed
- Output limit32K236K7.4×
- Input price$0.048/M$0.09/M▲ 86.9%
- Output price$0.193/M$0.3/M▲ 55.4%
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Context200K205K1×
- Output limit128K131K1×
via nearai
What changed
Removed from the model catalog
via hyper
What changed
- Input price$1.36/M$1.39/M▲ 2.1%
- Output price$4.27/M$4.36/M▲ 2.1%
- cache write$0.679/M$0.693/M▲ 2.1%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nearai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.755/M$0.757/M▲ 0.3%
- Output price$0.755/M$0.757/M▲ 0.3%
via amazon-bedrock
What changed
First observed in the model catalog
via llmgateway
What changed
- Input price$0.1/M$0.088/M▼ 12.0%
- Cache read$0.02/M$0.025/M▲ 25.0%
via llmgateway
What changed
- Input price$1.99/M$2.20/M▲ 10.6%
- Output price$6.16/M$6.50/M▲ 5.5%
- Cache read$0.4/M$0.45/M▲ 12.5%
via greenpt
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via openrouter
What changed
- nameLlama 3.1 70B InstructLlama-3.1-70B-Instruct
- knowledge2023-12-312023-12
via nearai
What changed
- Output limit33K8K−75%
via llmgateway-providers
What changed
- Input price$0.13/M$0.088/M▼ 32.3%
- Output price$0.4/M$0.25/M▼ 37.5%
- Cache read$0.024/M$0.025/M▲ 4.2%
via amazon-bedrock
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via venice
What changed
- Input price$0.05/M$0.25/M▲ 400.0%
- Output price$0.187/M$0.938/M▲ 400.0%
- Cache read$0.0050/M$0.025/M▲ 400.0%
via llmgateway
What changed
- Input price$0.55/M$0.8/M▲ 45.5%
- Output price$1.78/M$2.55/M▲ 42.9%
- Cache read$0.111/M$0.16/M▲ 44.1%
via amazon-bedrock
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit393K384K−2%
- Input price$0.579/M$1.05/M▲ 81.1%
- Output price$1.74/M$3.15/M▲ 81.1%
- Cache read$0.058/M$0.035/M▼ 39.6%
via nearai
What changed
Removed from the model catalog
via hyper
What changed
- Input price$0.178/M$0.18/M▲ 1.1%
- Output price$0.68/M$0.61/M▼ 10.3%
- cache write$0.089/M$0.09/M▲ 1.1%
via amazon-bedrock
What changed
- structured outputNoYesEnabled
via amazon-bedrock
What changed
First observed in the model catalog
via nearai
What changed
- modalities.inputtext,imagetext
via crossmodel
What changed
- Input price$0.075/M$0.15/M▲ 100.0%
- Output price$0.25/M$0.5/M▲ 100.0%
- Cache read$0.015/M$0.03/M▲ 100.0%
- cache write$0.075/M$0.15/M▲ 100.0%
via amazon-bedrock
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit128K131K1×
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$1.08/M$1.20/M▲ 11.1%
- Cache read$0.027/M$0.03/M▲ 11.1%
via nearai
What changed
- Output price$0/M$0.01/M
via amazon-bedrock
What changed
First observed in the model catalog
via cortecs
What changed
- namepixtral-large-2502Pixtral Large (25.02)
- family—pixtral
- release date2025-05-262025-04-08
- last updated2025-05-262025-04-08
via nano-gpt
What changed
Removed from the model catalog
via digitalocean
What changed
- Input price$2.85/M$2.55/M▼ 10.5%
- Output price$14.25/M$12.95/M▼ 9.1%
via amazon-bedrock
What changed
First observed in the model catalog
via greenpt
What changed
First observed in the model catalog
via nearai
What changed
- Output limit131K16K−88%
- Input price$0.85/M$1.40/M▲ 64.7%
- Output price$3.30/M$4.40/M▲ 33.3%
via edenai
What changed
- Input price$1.05/M$1.05/M▲ 0.3%
- Output price$1.05/M$1.05/M▲ 0.3%
via amazon-bedrock
What changed
First observed in the model catalog
via kilo
What changed
Removed from the model catalog
via vercel
What changed
Removed from the model catalog
via kilo
What changed
- Context164K128K−22%
- Output limit16K16K−2%
- Input price$0.32/M$0.257/M▼ 19.6%
- Output price$0.89/M$1.03/M▲ 15.6%
via nearai
What changed
- Context41K33K−20%
- Output price$0/M$0.01/M
via amazon-bedrock
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.174/M$0.175/M▲ 0.3%
- Output price$0.755/M$0.757/M▲ 0.3%
via nearai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.465/M$0.466/M▲ 0.3%
- Output price$0.929/M$0.932/M▲ 0.3%
via amazon-bedrock
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$1.30/M$1.40/M▲ 7.7%
- Output price$4/M$4.40/M▲ 10.0%
- Cache read$0.25/M$0.26/M▲ 4.0%
via amazon-bedrock
What changed
First observed in the model catalog
via kilo
What changed
- nameMeta: Llama 3.1 70B InstructLlama-3.1-70B-Instruct
- open weightsNoYesEnabled
- knowledge—2023-12
via amazon-bedrock
What changed
First observed in the model catalog
via kilo
What changed
- structured outputYesNoRemoved
via 302ai
What changed
- descriptionMiniMax model for chat, coding, office work, and agentic tasksOpen MiniMax flagship for coding agents, office automation, and complex environments
- family—minimax
- reasoningNoYesEnabled
- open weightsNoYesEnabled
- +2 more changes
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via requesty
What changed
- Output price$3/M$2.60/M▼ 13.3%
- Cache read$0.26/M$0.05/M▼ 80.8%
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via privatemode-ai
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via kilo
What changed
- Context1.02M1.05M1×
- Output limit384K393K1×
- Cache read$0.132/M$0.044/M▼ 66.7%
via amazon-bedrock
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.1%
- Output price$0.697/M$0.697/M▼ 0.1%
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.1/M$0.09/M▼ 10.0%
- Cache read$0.07/M—
via llmgateway-providers
What changed
- Input price$0.41/M$0.2/M▼ 51.2%
- Output price$2.50/M$2/M▼ 20.0%
- Cache read$0.08/M$0.05/M▼ 37.5%
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via edenai
What changed
- Input price$0.03/M$0.065/M▲ 116.7%
- Output price$0.1/M$0.18/M▲ 80.0%
via openrouter
What changed
- Output limit131K944K7.2×
- Input price$0.14/M$0.065/M▼ 53.6%
- Output price$0.28/M$0.18/M▼ 35.7%
- Cache read$0.028/M$0.016/M▼ 42.9%
via 302ai
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.02/M$0.03/M▲ 50.0%
- Output price$0.1/M$0.13/M▲ 30.0%
via kilo
What changed
- Input price$2.50/M$2.40/M▼ 4.0%
- Output price$14/M$12/M▼ 14.3%
- Cache read$0.29/M$0.24/M▼ 17.2%
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via openrouter
What changed
- Output limit944K128K−86%
- Input price$1.12/M$1.11/M▼ 0.6%
- Output price$3.52/M$3.50/M▼ 0.6%
- Cache read$0.208/M$0.207/M▼ 0.6%
via kilo
What changed
- Output limit384K944K2.5×
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via kilo
What changed
- Output limit262K131K−50%
via 302ai
What changed
First observed in the model catalog
via kilo
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video,audio
via 302ai
What changed
Removed from the model catalog
via kilo
What changed
- Output limit66K236K3.6×
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,audio
via vercel
What changed
Removed from the model catalog
via 302ai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$1.05/M$1.05/M▼ 0.1%
- Output price$1.05/M$1.05/M▼ 0.1%
via 302ai
What changed
Removed from the model catalog
via llmgateway
What changed
- Input price$0.41/M$0.2/M▼ 51.2%
- Output price$2.50/M$2/M▼ 20.0%
- Cache read$0.08/M$0.05/M▼ 37.5%
via 302ai
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit384K944K2.5×
- Input price$0.44/M$0.22/M▼ 50.0%
- Output price$1.32/M$0.66/M▼ 50.0%
- Cache read$0.014/M$0.0070/M▼ 50.0%
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via venice
What changed
First observed in the model catalog
via kilo
What changed
Removed from the model catalog
via 302ai
What changed
Removed from the model catalog
via zeldoc
What changed
- knowledge2025-01—
- release date2026-04-152026-08-26
- last updated2026-04-152026-08-26
via vercel
What changed
- Input price$0.7/M$1.40/M▲ 100.0%
- Output price$2.20/M$4.40/M▲ 100.0%
- Cache read$0.13/M$0.14/M▲ 7.7%
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via 302ai
What changed
- nameglm-5.1GLM-5.1
- descriptionFlagship GLM model for hybrid reasoning, coding, and agentic engineeringStrong GLM coding model for agentic engineering, terminals, and repository generation
- open weightsNoYesEnabled
- release date2026-04-102026-04-07
- +3 more changes
via nano-gpt
What changed
- cache write$0.075/M$0.042/M▼ 44.4%
via vercel
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via 302ai
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit262K131K−50%
via llmgateway
What changed
- Input price$0.22/M$0.6/M▲ 172.7%
- Output price$1.14/M$3.05/M▲ 168.2%
- Cache read$0.048/M$0.13/M▲ 170.8%
via requesty
What changed
- Input price$0.14/M$0.28/M▲ 100.0%
- Output price$0.28/M$0.56/M▲ 100.0%
via 302ai
What changed
First observed in the model catalog
via privatemode-ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.465/M$0.465/M▼ 0.1%
- Output price$0.93/M$0.929/M▼ 0.1%
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via openrouter
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via kilo
What changed
Removed from the model catalog
via kilo
What changed
- Output limit131K944K7.2×
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via cortecs
What changed
First observed in the model catalog
via requesty
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via hyper
What changed
- Input price$0.188/M$0.178/M▼ 5.3%
- Output price$0.7/M$0.68/M▼ 2.9%
- cache write$0.094/M$0.089/M▼ 5.3%
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via orcarouter
What changed
Removed from the model catalog
via 302ai
What changed
Removed from the model catalog
via requesty
What changed
- Input price$2.25/M$3/M▲ 33.3%
- Output price$11.25/M$15/M▲ 33.3%
- Cache read$0.225/M$0.45/M▲ 100.0%
via kilo
What changed
- Output limit236K16K−93%
via requesty
What changed
- Output price$0.5/M$0.6/M▲ 20.0%
via 302ai
What changed
First observed in the model catalog
via requesty
What changed
- Output price$0.5/M$0.6/M▲ 20.0%
via kilo
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.11/M$0.106/M▼ 3.6%
- Output price$0.408/M$0.368/M▼ 9.8%
- cache write$0.055/M$0.053/M▼ 3.6%
via kilo
What changed
- Context1.05M1.02M−2%
- Output limit944K128K−86%
- Input price$1.12/M$1.11/M▼ 0.6%
- Output price$3.52/M$3.50/M▼ 0.6%
- +1 more changes
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
- descriptionMiniMax model for chat, coding, office work, and agentic tasksEarlier MiniMax agent model for practical coding and productivity tasks
- family—minimax
- reasoningNoYesEnabled
- open weightsNoYesEnabled
- +3 more changes
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via openrouter
What changed
- Output limit131K16K−88%
- Input price$0.55/M$0.43/M▼ 21.8%
- Output price$2.20/M$1.75/M▼ 20.5%
- Cache read$0.11/M$0.08/M▼ 27.3%
via 302ai
What changed
Removed from the model catalog
via hyper
What changed
- Input price$1.33/M$1.36/M▲ 2.0%
- Output price$4.31/M$4.27/M▼ 1.0%
- cache write$0.666/M$0.679/M▲ 2.0%
via nano-gpt
What changed
- modalities.inputtext,image,audio,pdftext,image,video,audio,pdf
via amazon-bedrock
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.4/M$0.55/M▲ 37.5%
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
- cache write$0.075/M$0.042/M▼ 44.4%
via openrouter
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.47/M$0.458/M▼ 2.6%
- Output price$1.76/M$1.71/M▼ 2.7%
- cache write$0.235/M$0.229/M▼ 2.6%
via requesty
What changed
- Input price$0.14/M$0.28/M▲ 100.0%
- Output price$0.28/M$0.56/M▲ 100.0%
via openrouter
What changed
- structured outputYesNoRemoved
- Context1M262K−74%
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,audio
via nano-gpt
What changed
- tool callYesNoRemoved
via openrouter
What changed
- Input price$0.1/M$0.06/M▼ 40.0%
- Output price$0.15/M$0.25/M▲ 66.7%
- Cache read$0.05/M$0.015/M▼ 70.0%
via openrouter
What changed
- Output limit384K393K1×
- Input price$1.05/M$0.579/M▼ 44.8%
- Output price$3.15/M$1.74/M▼ 44.8%
- Cache read$0.035/M$0.058/M▲ 65.7%
via 302ai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.1%
- Output price$0.755/M$0.755/M▼ 0.1%
via nano-gpt
What changed
- Input price$0.045/M$0.054/M▲ 20.0%
- Output price$0.177/M$0.212/M▲ 20.0%
- Cache read$0.028/M$0.034/M▲ 20.0%
via 302ai
What changed
Removed from the model catalog
via crossmodel
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.039/M$0.037/M▼ 5.1%
- Output price$0.1/M$0.17/M▲ 70.0%
via llmgateway-providers
What changed
Removed from the model catalog
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- tool callYesNoRemoved
via nano-gpt
What changed
- modalities.inputtext,image,videotext,image,audio,video
via 302ai
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
- modalities.inputtext,image,videotext,image,audio,video
via 302ai
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$2.50/M$2.40/M▼ 4.0%
- Output price$14/M$12/M▼ 14.3%
- Cache read$0.29/M$0.24/M▼ 17.2%
via 302ai
What changed
First observed in the model catalog
via kilo
What changed
- Output limit131K944K7.2×
- Input price$0.071/M$0.075/M▲ 5.3%
- Output price$0.237/M$0.25/M▲ 5.3%
- Cache read$0.014/M$0.015/M▲ 5.3%
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- attachmentNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via digitalocean
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via requesty
What changed
- Output price$3/M$2.60/M▼ 13.3%
- Cache read$0.26/M$0.05/M▼ 80.8%
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit131K944K7.2×
- Input price$0.071/M$0.075/M▲ 5.3%
- Output price$0.237/M$0.25/M▲ 5.3%
- Cache read$0.014/M$0.015/M▲ 5.3%
via nano-gpt
What changed
- modalities.inputtext,imagetext,image,video
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via 302ai
What changed
First observed in the model catalog
via kilo
What changed
- Context205K198K−3%
- Output limit131K16K−88%
- Input price$0.55/M$0.43/M▼ 21.8%
- Output price$2.20/M$1.75/M▼ 20.5%
- +1 more changes
via nano-gpt
What changed
- Input price$0.07/M$0.35/M▲ 400.0%
- Output price$0.21/M$1.40/M▲ 566.7%
- Cache read$0.014/M$0.175/M▲ 1150.0%
via nano-gpt
What changed
- modalities.inputtext,image,audiotext,image,video,audio
via 302ai
What changed
Removed from the model catalog
via orcarouter
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.755/M$0.755/M▼ 0.1%
- Output price$0.755/M$0.755/M▼ 0.1%
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit66K236K3.6×
- Input price$0.39/M$0.55/M▲ 41.0%
- Output price$2.34/M$3.50/M▲ 49.6%
- Cache read—$0.225/M
via nano-gpt
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.14/M$0.28/M▲ 100.0%
- Output price$0.28/M$0.56/M▲ 100.0%
via 302ai
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via vercel
What changed
Removed from the model catalog
via openrouter
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via 302ai
What changed
Removed from the model catalog
via kilo
What changed
First observed in the model catalog
via requesty
What changed
- Input price$2.25/M$3/M▲ 33.3%
- Output price$11.25/M$15/M▲ 33.3%
- Cache read$0.225/M$0.45/M▲ 100.0%
via 302ai
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via 302ai
What changed
First observed in the model catalog
via 302ai
What changed
First observed in the model catalog
via nano-gpt
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image,video
via crof
What changed
- open weightsNoYesEnabled
via aihubmix
What changed
- open weightsNoYesEnabled
via requesty
What changed
- open weightsNoYesEnabled
via zhipuai
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$2.55/M$2.50/M▼ 2.0%
- Output price$12.75/M$14/M▲ 9.8%
- Cache read$0.256/M$0.29/M▲ 13.3%
via cortecs
What changed
Removed from the model catalog
via hyper
What changed
- Input price$0.484/M$0.47/M▼ 2.9%
- Output price$1.85/M$1.76/M▼ 5.0%
- cache write$0.242/M$0.235/M▼ 2.9%
via kilo
What changed
- Input price$2.55/M$2.50/M▼ 2.0%
- Output price$12.75/M$14/M▲ 9.8%
- Cache read$0.256/M$0.29/M▲ 13.3%
via nano-gpt
What changed
- open weightsNoYesEnabled
via hyper
What changed
- Input price$0.106/M$0.11/M▲ 3.8%
- Output price$0.368/M$0.408/M▲ 10.9%
- cache write$0.053/M$0.055/M▲ 3.8%
via hyper
What changed
- open weightsNoYesEnabled
via vercel
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Context41K131K3.2×
- Output limit16K8K−50%
via openrouter
What changed
Removed from the model catalog
via zhipuai-coding-plan
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Context203K131K−35%
- Output limit16K118K7.2×
- Input price$0.06/M$0.06/M▲ 0.8%
- Cache read$0.01/M—
via openrouter
What changed
- Output limit236K944K4×
- Input price$1.17/M$1.12/M▼ 4.3%
- Output price$3.96/M$3.52/M▼ 11.1%
- Cache read$0.234/M$0.208/M▼ 11.1%
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Output limit145K33K−77%
- Input price$0.55/M$0.25/M▼ 54.5%
- Output price$1.65/M$0.95/M▼ 42.4%
- Cache read$0.55/M$0.13/M▼ 76.4%
via cortecs
What changed
- Input price$1.40/M$1.11/M▼ 20.4%
- Output price$4.40/M$3.90/M▼ 11.4%
- Cache read$0.26/M$0.279/M▲ 7.3%
via synthetic
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$1.04/M$0.955/M▼ 7.7%
- Output price$2.07/M$1.91/M▼ 7.7%
- Cache read$0.086/M$0.08/M▼ 7.7%
via openrouter
What changed
- Output limit944K131K−86%
- Input price$0.075/M$0.071/M▼ 5.0%
- Output price$0.25/M$0.237/M▼ 5.0%
- Cache read$0.015/M$0.014/M▼ 5.0%
via llmgateway-providers
What changed
- attachmentNoYesEnabled
- open weightsNoYesEnabled
- modalities.inputtexttext,image,video,pdf
via kilo
What changed
- Output limit944K131K−86%
- Input price$0.075/M$0.071/M▼ 5.0%
- Output price$0.25/M$0.237/M▼ 5.0%
- Cache read$0.015/M$0.014/M▼ 5.0%
via nano-gpt
What changed
First observed in the model catalog
via modal
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$0.25/M$0.29/M▲ 16.0%
- Output price$1/M$1.14/M▲ 14.0%
- Cache read—$0.11/M
via opencode-go
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$0.66/M$0.71/M▲ 7.6%
- Output price$3.40/M$3.50/M▲ 2.9%
- Cache read$0.18/M$0.15/M▼ 16.7%
via nano-gpt
What changed
- Output price$4.60/M$4.40/M▼ 4.3%
- Cache read$0.5/M$0.7/M▲ 40.0%
via kilo
What changed
- Context161K164K1×
- Output limit145K33K−77%
via nebius
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Input price$0.25/M$0.29/M▲ 16.0%
- Output price$1/M$1.14/M▲ 14.0%
- Cache read—$0.11/M
via hyper
What changed
- Input price$1.26/M$1.33/M▲ 5.5%
- Output price$4.13/M$4.31/M▲ 4.4%
- cache write$0.631/M$0.666/M▲ 5.5%
via cortecs
What changed
Removed from the model catalog
via vivgrid
What changed
First observed in the model catalog
via tokengo
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Output limit100K236K2.4×
via cortecs
What changed
- open weightsNoYesEnabled
via fireworks-ai
What changed
- last updated2026-09-042026-09-07
- Context1M1.05M1×
- Output limit131K262K2×
via requesty
What changed
- open weightsNoYesEnabled
via baseten
What changed
- open weightsNoYesEnabled
via vercel
What changed
Removed from the model catalog
via kilo
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via nano-gpt
What changed
Removed from the model catalog
via vivgrid
What changed
First observed in the model catalog
via kenari
What changed
- open weightsNoYesEnabled
via cortecs
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit944K393K−58%
- Input price$0.045/M$0.05/M▲ 11.1%
- Output price$0.09/M$0.16/M▲ 77.8%
- Cache read$0.0090/M$0.013/M▲ 44.4%
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via orcarouter
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.35/M$0.07/M▼ 80.0%
- Output price$1.40/M$0.21/M▼ 85.0%
- Cache read$0.175/M$0.014/M▼ 92.0%
via openrouter
What changed
- Output limit16K131K8×
- Input price$0.43/M$0.55/M▲ 27.9%
- Output price$1.75/M$2.20/M▲ 25.7%
- Cache read$0.08/M$0.11/M▲ 37.5%
via fireworks-ai
What changed
- open weightsNoYesEnabled
- last updated2026-09-042026-09-07
- Context1M1.05M1×
via kilo
What changed
- Context198K205K1×
- Output limit16K131K8×
- Input price$0.43/M$0.55/M▲ 27.9%
- Output price$1.75/M$2.20/M▲ 25.7%
- +1 more changes
via vancine
What changed
- open weightsNoYesEnabled
via deepinfra
What changed
- status—deprecated
via above
What changed
- open weightsNoYesEnabled
via opencode
What changed
- open weightsNoYesEnabled
via cortecs
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via deepinfra
What changed
- open weightsNoYesEnabled
via zai
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$1.12/M$1.05/M▼ 6.4%
- Output price$3.36/M$3.15/M▼ 6.4%
- Cache read$0.037/M$0.035/M▼ 6.4%
via openrouter
What changed
- Output limit16K8K−50%
- Input price$0.12/M$0.228/M▲ 89.6%
- Output price$0.24/M$0.91/M▲ 279.2%
via kilo
What changed
Removed from the model catalog
via merge-gateway
What changed
- open weightsNoYesEnabled
via runinfra
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$0.09/M$0.089/M▼ 1.4%
- Output price$0.18/M$0.177/M▼ 1.4%
- Cache read$0.018/M$0.018/M▼ 1.4%
via kilo
What changed
- structured outputNoYesEnabled
via openrouter
What changed
- open weightsNoYesEnabled
via amazon-bedrock
What changed
First observed in the model catalog
via kilo
What changed
- Output limit944K393K−58%
- Input price$0.045/M$0.05/M▲ 11.1%
- Output price$0.09/M$0.16/M▲ 77.8%
- Cache read$0.0090/M$0.013/M▲ 44.4%
via tinfoil
What changed
- open weightsNoYesEnabled
via digitalocean
What changed
- open weightsNoYesEnabled
via edenai
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Context262K1.05M4×
- Output limit236K944K4×
- Input price$1.17/M$1.12/M▼ 4.3%
- Output price$3.96/M$3.52/M▼ 11.1%
- +1 more changes
via edenai
What changed
First observed in the model catalog
via edenai
What changed
- Input price$2/M$1.50/M▼ 25.0%
- Output price$5/M$7.50/M▲ 50.0%
- Cache read—$0.15/M
via crossmodel
What changed
- open weightsNoYesEnabled
via berget
What changed
- open weightsNoYesEnabled
via vivgrid
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit100K236K2.4×
- Cache read$0.15/M—
via ofox
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$0.55/M$0.4/M▼ 27.3%
via llmgateway
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Output price$4.60/M$4.40/M▼ 4.3%
- Cache read$0.5/M$0.7/M▲ 40.0%
via openrouter
What changed
- Output limit16K118K7.2×
- Input price$0.06/M$0.06/M▲ 0.8%
- Cache read$0.01/M—
via vivgrid
What changed
First observed in the model catalog
via kilo
What changed
- open weightsNoYesEnabled
via hyper
What changed
- Input price$0.514/M$0.558/M▲ 8.6%
- Output price$2.75/M$2.94/M▲ 6.5%
- cache write$0.257/M$0.279/M▲ 8.6%
via scnet-token-plan
What changed
- open weightsNoYesEnabled
via vercel
What changed
Removed from the model catalog
via empiriolabs
What changed
- open weightsNoYesEnabled
via amazon-bedrock
What changed
First observed in the model catalog
via vivgrid
What changed
- open weightsNoYesEnabled
via openrouter
What changed
Removed from the model catalog
via llmgateway-providers
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via vivgrid
What changed
First observed in the model catalog
via togetherai
What changed
- open weightsNoYesEnabled
via gitlab
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.14/M$0.132/M▼ 5.7%
- Output price$0.58/M$0.528/M▼ 9.0%
- Cache read$0.035/M$0.033/M▼ 5.7%
via nano-gpt
What changed
- open weightsNoYesEnabled
via zai-coding-plan
What changed
- open weightsNoYesEnabled
via fireworks-ai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via edenai
What changed
Removed from the model catalog
via kilo
What changed
Removed from the model catalog
via openrouter
What changed
- structured outputNoYesEnabled
via nan
What changed
- open weightsNoYesEnabled
via neon
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Context203K198K−2%
- Output limit131K16K−88%
- Input price$0.5/M$0.43/M▼ 14.0%
- Output price$2/M$1.75/M▼ 12.5%
- +1 more changes
via edenai
What changed
- nameClaude Fable 5.1Claude Fable Latest (Claude Fable 5.1)
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (FlexAI)
via edenai
What changed
- nameNemotron 3 Ultra 550B A55BNemotron 3 Ultra 550B A55B (Nebius)
via kilo
What changed
- Output limit236K66K−72%
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Groq)
via edenai
What changed
- nameGemini 3.8 FlashGemini 3.8 Flash (Vertex AI)
via edenai
What changed
- nameGemini 3.7 Flash (US)Gemini 3.7 Flash (Vertex AI, US)
via fastrouter
What changed
- nameNano Banana 2Nano Banana 2 Preview
via moonshotai
What changed
Removed from the model catalog
via deepinfra
What changed
- Input price$0.08/M$0.06/M▼ 25.0%
- Cache read$0.016/M$0.015/M▼ 6.3%
via venice
What changed
- Input price$3.13/M$2.50/M▼ 20.0%
- Output price$18.75/M$15/M▼ 20.0%
- Cache read$0.313/M$0.25/M▼ 20.0%
- cache write$3.91/M$3.13/M▼ 20.0%
- +1 more changes
via edenai
What changed
- nameNemotron 3 Super 120B A12BNemotron 3 Super 120B A12B (Nebius)
via edenai
What changed
- nameKimi K2.5Kimi K2.5 (Amazon Bedrock)
via edenai
What changed
- nameInkling SmallInkling Small (Deep Infra)
via edenai
What changed
- nameNano Banana 2 LiteNano Banana 2 Lite (Vertex AI)
via llmgateway-providers
What changed
- knowledge—2026-04-30
via nano-gpt
What changed
- knowledge—2026-04-30
via edenai
What changed
- nameGPT-6 AstraGPT Latest (GPT-6 Astra)
- knowledge—2026-04-30
via edenai
What changed
- nameGPT-5.2 CodexGPT-5.2 Codex (Azure)
via abacus
What changed
- nameNano Banana ProNano Banana Pro Preview
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (OVHcloud)
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (Together AI)
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (Databricks)
via edenai
What changed
- nameGPT-5.4 miniGPT Mini Latest (GPT-5.4 mini)
via edenai
What changed
- nameMuse Glimmer 30BMuse Glimmer 30B (Together AI)
via edenai
What changed
- nameDeepSeek-R1DeepSeek-R1 (Deep Infra)
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (TensorX)
via sensenova
What changed
First observed in the model catalog
via edenai
What changed
- nameNano BananaNano Banana (Vertex AI)
via edenai
What changed
- nameGLM-4.7-Flash (US)GLM-4.7-Flash (Amazon Bedrock, US)
via edenai
What changed
- nameLlama-Guard-3-8BLlama-Guard-3-8B (Cloudflare)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Deep Infra)
via edenai
What changed
- nameGemini 3.7 FlashGemini 3.7 Flash (Vertex AI)
via edenai
What changed
- nameStep 3.7 FlashStep 3.7 Flash (Deep Infra)
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (Alibaba)
via edenai
What changed
- nameGemini 3.5 Flash (US)Gemini 3.5 Flash (Vertex AI, US)
via edenai
What changed
- nameGPT-5.1 Codex miniGPT-5.1 Codex mini (Azure)
via edenai
What changed
- nameLlama-3.3-70B-InstructLlama-3.3-70B-Instruct (Nebius)
via edenai
What changed
- nameGPT OSS 20B (EU)GPT OSS 20B (Databricks, EU)
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (Together AI)
via openrouter
What changed
- Output limit131K944K7.2×
- Input price$0.05/M$0.045/M▼ 10.0%
- Output price$0.1/M$0.09/M▼ 10.0%
- Cache read$0.0100/M$0.0090/M▼ 10.0%
via moonshotai-cn
What changed
Removed from the model catalog
via edenai
What changed
- nameLlama-Guard-3-8BLlama-Guard-3-8B (Deep Infra)
via edenai
What changed
- knowledge—2026-04-30
via openrouter
What changed
- Input price$0.779/M$1.04/M▲ 32.9%
- Output price$1.56/M$2.07/M▲ 32.9%
- Cache read$0.065/M$0.086/M▲ 32.9%
via crof
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via edenai
What changed
- nameSeed 2.0 MiniSeed 2.0 Mini (Deep Infra)
via edenai
What changed
- nameGLM-4.7-FlashGLM-4.7-Flash (Amazon Bedrock)
via nano-gpt
What changed
- Context1M400K−60%
- Output limit33K128K3.9×
- Input limit1M400K−60%
via neon
What changed
First observed in the model catalog
via edenai
What changed
- nameGemma-SEA-LION-v4-27B-ITGemma-SEA-LION-v4-27B-IT (Cloudflare)
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Alibaba)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Nebius)
via edenai
What changed
- nameGemini 3.6 Flash (US)Gemini 3.6 Flash (Vertex AI, US)
via edenai
What changed
- nameGLM-4.7-FlashGLM-4.7-Flash (Deep Infra)
via sensenova
What changed
First observed in the model catalog
via edenai
What changed
- nameNano Banana 2Nano Banana 2 Preview
via edenai
What changed
- nameHy3Hy3 (Deep Infra)
via abacus
What changed
- nameNano Banana 2Nano Banana 2 Preview
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (Databricks)
via kilo
What changed
- Context131K262K2×
- Output limit33K236K7.2×
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (FlexAI)
via edenai
What changed
- nameGemini 3.5 FlashGemini 3.5 Flash (Vertex AI)
via edenai
What changed
- nameInklingInkling (Fireworks AI)
via edenai
What changed
- nameSeed 2.0 CodeSeed 2.0 Code (Deep Infra)
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (Cloudflare)
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Scaleway)
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (OVHcloud)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Databricks)
via edenai
What changed
- nameGPT OSS 120B (EU)GPT OSS 120B (Databricks, EU)
via edenai
What changed
- nameNano Banana 2Nano Banana 2 (Vertex AI)
via edenai
What changed
- nameInkling SmallInkling Small (Together AI)
via edenai
What changed
- nameGPT-5.1 CodexGPT-5.1 Codex (Azure)
via moonshotai
What changed
Removed from the model catalog
via edenai
What changed
- nameGLM-4.7-FlashGLM-4.7-Flash (Cloudflare)
via edenai
What changed
- nameLlama-3.2-11B-Vision-InstructLlama-3.2-11B-Vision-Instruct (Deep Infra)
via edenai
What changed
- nameInklingInkling (Deep Infra)
via edenai
What changed
- nameGemini 3.5 Flash (EU)Gemini 3.5 Flash (Vertex AI, EU)
via edenai
What changed
- nameKimi K2.5Kimi K2.5 (Deep Infra)
via moonshotai-cn
What changed
Removed from the model catalog
via tinfoil
What changed
- reasoningYesNoRemoved
via edenai
What changed
- nameMuse Glimmer 30BMuse Glimmer 30B (FlexAI)
via azure-cognitive-services
What changed
- statusbeta—
via edenai
What changed
- nameNano Banana ProNano Banana Pro Preview
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Databricks)
via edenai
What changed
- nameGrok 4.6Grok Latest (Grok 4.6)
via azure-cognitive-services
What changed
- statusbeta—
via moonshotai
What changed
Removed from the model catalog
via moonshotai-cn
What changed
Removed from the model catalog
via edenai
What changed
- nameGemini 3.5 Flash LiteGemini 3.5 Flash Lite (Vertex AI)
via venice
What changed
- Input price$6.25/M$2.50/M▼ 60.0%
- Output price$37.50/M$12.50/M▼ 66.7%
- Cache read$0.625/M$0.25/M▼ 60.0%
- cache write$7.81/M$3.13/M▼ 60.0%
- +1 more changes
via azure
What changed
- statusbeta—
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (Fireworks AI)
via kilo
What changed
- Output limit262K944K3.6×
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Deep Infra)
- Input price$0.08/M$0.06/M▼ 25.0%
- Cache read$0.016/M$0.015/M▼ 6.3%
via edenai
What changed
- nameKimi K2.5Kimi K2.5 (TensorX)
via edenai
What changed
- nameGemini 3.6 Flash (EU)Gemini 3.6 Flash (Vertex AI, EU)
via azure
What changed
First observed in the model catalog
via openrouter
What changed
- knowledge—2026-04-30
via azure-cognitive-services
What changed
- statusbeta—
via moonshotai-cn
What changed
Removed from the model catalog
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Together AI)
via azure
What changed
- statusbeta—
via kilo
What changed
- knowledge—2026-04-30
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (TensorX)
via nano-gpt
What changed
- Context1M1.05M1.1×
- Input limit1M1.05M1.1×
via edenai
What changed
- nameGemini 3.1 Pro PreviewGemini Pro Latest (Gemini 3.1 Pro Preview, Vertex AI)
via edenai
What changed
- nameMuse Glimmer 30BMuse Glimmer 30B (Deep Infra)
via edenai
What changed
- nameQwen2.5-Coder-32B-InstructQwen2.5-Coder-32B-Instruct (Cloudflare)
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (Cloudflare)
via openrouter
What changed
- Output limit33K236K7.2×
via edenai
What changed
- nameNemotron 3 Ultra 550B A55BNemotron 3 Ultra 550B A55B (Deep Infra)
via edenai
What changed
- nameDeepSeek-V3DeepSeek-V3 (Deep Infra)
via azure
What changed
- statusbeta—
via openrouter
What changed
- nameNano Banana ProNano Banana Pro Preview
via cerebras
What changed
First observed in the model catalog
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (IONOS)
via edenai
What changed
- nameGemini 3.6 FlashGemini 3.6 Flash (Vertex AI)
via edenai
What changed
- nameGemini 3.1 Pro PreviewGemini 3.1 Pro Preview (Vertex AI)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Fireworks AI)
via edenai
What changed
- nameGemini 3.8 Flash (EU)Gemini 3.8 Flash (Vertex AI, EU)
via nano-gpt
What changed
- Context922K1.05M1.1×
- Input limit922K1.05M1.1×
via neon
What changed
First observed in the model catalog
via merge-gateway
What changed
- knowledge—2026-04-30
via edenai
What changed
- nameLlama-3.3-70B-InstructLlama-3.3-70B-Instruct (Deep Infra)
via edenai
What changed
- nameGemini 3.5 Flash Lite (US)Gemini 3.5 Flash Lite (Vertex AI, US)
via github-copilot
What changed
- knowledge—2026-04-30
via edenai
What changed
- nameClaude Sonnet 5Claude Sonnet Latest (Claude Sonnet 5)
via crossmodel
What changed
- knowledge—2026-04-30
via ofox
What changed
First observed in the model catalog
via venice
What changed
- knowledge—2026-04-30
via azure
What changed
- statusbeta—
via neon
What changed
First observed in the model catalog
via edenai
What changed
- nameMuse Glimmer 30BMuse Glimmer 30B (Fireworks AI)
via edenai
What changed
- nameLlama-3.3-70B-InstructLlama-3.3-70B-Instruct (IONOS)
via edenai
What changed
- nameLlama-3.3-70B-InstructLlama-3.3-70B-Instruct (Scaleway)
via edenai
What changed
- nameGemini 3.1 Flash LiteGemini 3.1 Flash Lite (Vertex AI)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Scaleway)
via edenai
What changed
- nameStep 3.5 FlashStep 3.5 Flash (Deep Infra)
via edenai
What changed
- nameNemotron 3 Super 120B A12BNemotron 3 Super 120B A12B (FlexAI)
via openrouter
What changed
- Input price$0.082/M$0.09/M▲ 9.0%
- Output price$0.165/M$0.18/M▲ 9.0%
- Cache read$0.016/M$0.018/M▲ 9.0%
via edenai
What changed
First observed in the model catalog
via edenai
What changed
- nameGemini 3.8 FlashGemini Flash Latest (Gemini 3.8 Flash, Vertex AI)
via edenai
What changed
- nameInklingInkling (Together AI)
via nano-gpt
What changed
First observed in the model catalog
via venice
What changed
- Input price$6.25/M$2.50/M▼ 60.0%
- Output price$37.50/M$12.50/M▼ 66.7%
- Cache read$0.625/M$0.25/M▼ 60.0%
- cache write$7.81/M$3.13/M▼ 60.0%
- +1 more changes
via kilo
What changed
- nameNano Banana ProNano Banana Pro Preview
via venice
What changed
- Input price$3.13/M$2.50/M▼ 20.0%
- Output price$18.75/M$15/M▼ 20.0%
- Cache read$0.313/M$0.25/M▼ 20.0%
- cache write$3.91/M$3.13/M▼ 20.0%
- +1 more changes
via edenai
What changed
- nameClaude Opus 5Claude Opus Latest (Claude Opus 5)
via openrouter
What changed
- Output limit262K944K3.6×
- Cache read$0.14/M$0.26/M▲ 85.7%
via cerebras
What changed
Removed from the model catalog
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (Groq)
via edenai
What changed
- nameDeepSeek V4 Pro 0813DeepSeek V4 Pro 0813 (Deep Infra)
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via openrouter
What changed
- Output limit944K131K−86%
- Input price$0.065/M$0.14/M▲ 115.4%
- Output price$0.18/M$0.28/M▲ 55.6%
- Cache read$0.016/M$0.028/M▲ 75.0%
via opencode
What changed
- knowledge—2026-04-30
via edenai
What changed
- nameNano Banana ProNano Banana Pro (Vertex AI)
via edenai
What changed
- nameGemini 3.1 Pro PreviewGemini Pro Latest (Gemini 3.1 Pro Preview)
via edenai
What changed
- nameKimi K2 ThinkingKimi K2 Thinking (Amazon Bedrock)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Together AI)
via edenai
What changed
- nameGemini 3.5 Flash Lite (EU)Gemini 3.5 Flash Lite (Vertex AI, EU)
via fastrouter
What changed
- nameNano Banana ProNano Banana Pro Preview
via moonshotai-cn
What changed
Removed from the model catalog
via edenai
What changed
- nameDeepSeek V3 0324DeepSeek V3 0324 (Deep Infra)
via azure-cognitive-services
What changed
- statusbeta—
via edenai
What changed
- nameGemini 3 Flash PreviewGemini 3 Flash Preview (Vertex AI)
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (Deep Infra)
via llmgateway-providers
What changed
Removed from the model catalog
via openai
What changed
- knowledge—2026-04-30
via azure
What changed
- statusbeta—
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Cloudflare)
via edenai
What changed
- nameStep 3.7 FlashStep 3.7 Flash (FlexAI)
via openrouter
What changed
Removed from the model catalog
via moonshotai
What changed
Removed from the model catalog
via edenai
What changed
- nameGPT OSS 20BGPT OSS 20B (FlexAI)
via edenai
What changed
- nameGemini 3.7 Flash (EU)Gemini 3.7 Flash (Vertex AI, EU)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Cloudflare)
via azure
What changed
- statusbeta—
via edenai
What changed
- nameGemini 3.1 Flash Lite (US)Gemini 3.1 Flash Lite (Vertex AI, US)
via edenai
What changed
- nameGPT-5.5 ProGPT Pro Latest (GPT-5.5 Pro)
via openrouter
What changed
- nameNano Banana 2Nano Banana 2 Preview
via edenai
What changed
- nameNemotron 3 Nano 30B A3BNemotron 3 Nano 30B A3B (Deep Infra)
via vercel
What changed
- knowledge—2026-04-30
via openrouter
What changed
- Output limit131K16K−88%
- Input price$0.5/M$0.43/M▼ 14.0%
- Output price$2/M$1.75/M▼ 12.5%
- Cache read$0.1/M$0.08/M▼ 20.0%
via edenai
What changed
- nameGemini 3.8 FlashGemini Flash Latest (Gemini 3.8 Flash)
via edenai
What changed
- nameGPT OSS 120BGPT OSS 120B (Cerebras)
via openrouter
What changed
- Output limit236K66K−72%
- Input price$0.55/M$0.39/M▼ 29.1%
- Output price$3.50/M$2.34/M▼ 33.1%
- Cache read$0.225/M—
via kilo
What changed
- nameNano Banana 2Nano Banana 2 Preview
via kilo
What changed
- Output limit131K944K7.2×
- Input price$0.05/M$0.045/M▼ 10.0%
- Output price$0.1/M$0.09/M▼ 10.0%
- Cache read$0.0100/M$0.0090/M▼ 10.0%
via edenai
What changed
- nameGPT-5.1 Codex MaxGPT-5.1 Codex Max (Azure)
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Nebius)
via edenai
What changed
- nameDeepSeek V4 Flash 0731DeepSeek V4 Flash 0731 (Fireworks AI)
via edenai
What changed
- nameLlama 3.1 Nemotron 70B InstructLlama 3.1 Nemotron 70B Instruct (Deep Infra)
via edenai
What changed
- nameGemini 3.8 Flash (US)Gemini 3.8 Flash (Vertex AI, US)
via neon
What changed
First observed in the model catalog
via kilo
What changed
- Output limit944K131K−86%
via moonshotai
What changed
Removed from the model catalog
via moonshotai
What changed
Removed from the model catalog
via moonshotai-cn
What changed
Removed from the model catalog
via edenai
What changed
- nameGemini 3.1 Flash Lite (EU)Gemini 3.1 Flash Lite (Vertex AI, EU)
via kilo
What changed
- Input price$1.15/M$1.17/M▲ 1.7%
- Output price$3.50/M$3.96/M▲ 13.1%
- Cache read$0.1/M$0.234/M▲ 134.0%
via kilo
What changed
- Context33K1.02M31.3×
- Output limit26K819K31.3×
via openrouter
What changed
- Input price$0.55/M$0.5/M▼ 9.1%
- Output price$2.20/M$2/M▼ 9.1%
- Cache read$0.11/M$0.1/M▼ 9.1%
via openrouter
What changed
- Input price$1.15/M$1.17/M▲ 1.7%
- Output price$3.50/M$3.96/M▲ 13.1%
- Cache read$0.1/M$0.234/M▲ 134.0%
via openrouter
What changed
- Input price$0.08/M$0.313/M▲ 290.6%
- Output price$0.75/M$1.25/M▲ 66.7%
- Cache read—$0.156/M
via venice
What changed
- Input price$1.25/M$0.25/M▼ 80.0%
- Output price$7.50/M$1.50/M▼ 80.0%
- Cache read$0.125/M$0.025/M▼ 80.0%
- cache write$1.56/M$0.313/M▼ 80.0%
- +1 more changes
via venice
What changed
- Input price$0.267/M$0.25/M▼ 6.3%
- Output price$1.60/M$1.50/M▼ 6.3%
- Cache read$0.027/M$0.025/M▼ 6.3%
- cache write$0.333/M$0.313/M▼ 6.3%
- +1 more changes
via kilo
What changed
- Context205K203K−1%
- Input price$0.55/M$0.5/M▼ 9.1%
- Output price$2.20/M$2/M▼ 9.1%
- Cache read$0.11/M$0.1/M▼ 9.1%
via nano-gpt
What changed
- Input price$2/M$10/M▲ 400.0%
- Output price$10/M$50/M▲ 400.0%
- Cache read$0.2/M$1/M▲ 400.0%
- cache write$2.50/M$12.50/M▲ 400.0%
via hyper
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.579/M$1.12/M▲ 93.4%
- Output price$1.74/M$3.36/M▲ 93.4%
- Cache read$0.019/M$0.037/M▲ 93.4%
via kilo
What changed
- Context262K256K−2%
via openrouter
What changed
- Output limit26K819K31.3×
via merge-gateway
What changed
- Cache read$0.03/M$0.0030/M▼ 90.0%
via openrouter
What changed
- Input price$0.084/M$0.082/M▼ 2.3%
- Output price$0.169/M$0.165/M▼ 2.3%
- Cache read$0.017/M$0.016/M▼ 2.3%
via openrouter
What changed
- Input price$0.85/M$0.779/M▼ 8.4%
- Output price$1.70/M$1.56/M▼ 8.4%
- Cache read$0.071/M$0.065/M▼ 8.4%
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via opencode
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via merge-gateway
What changed
- attachmentYesNoRemoved
- structured outputNoYesEnabled
- modalities.inputtext,imagetext
- Context256K131K−49%
- +1 more changes
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via nano-gpt
What changed
First observed in the model catalog
via baseten
What changed
First observed in the model catalog
via kilo
What changed
- Output limit236K16K−93%
via nano-gpt
What changed
- Context131K32K−76%
- Input limit131K32K−76%
via merge-gateway
What changed
- structured outputNoYesEnabled
via cline-pass
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via fireworks-ai
What changed
- last updated2026-08-282026-09-04
via edenai
What changed
- Input price$1.05/M$1.05/M▲ 0.1%
- Output price$1.05/M$1.05/M▲ 0.1%
via xai
What changed
- Context8K16K2×
via openrouter
What changed
- Input price$0.089/M$0.084/M▼ 4.7%
- Output price$0.177/M$0.169/M▼ 4.7%
- Cache read$0.018/M$0.017/M▼ 4.7%
via crusoe
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via github-copilot
What changed
First observed in the model catalog
via github-copilot
What changed
Removed from the model catalog
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via kilo
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- structured outputNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- +1 more changes
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.25/M$0.08/M▼ 68.0%
- Output price$1.25/M$0.75/M▼ 40.0%
- Cache read$0.25/M—
via nano-gpt
What changed
First observed in the model catalog
via xai
What changed
- Context8K16K2×
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via xai
What changed
- Context8K64K8×
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via openai
What changed
First observed in the model catalog
via scnet-token-plan
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via github-copilot
What changed
Removed from the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via nano-gpt
What changed
- Context128K32K−75%
- Input limit128K32K−75%
via vercel
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via kilo
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via kilo
What changed
- Output limit236K16K−93%
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via github-copilot
What changed
First observed in the model catalog
via github-copilot
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$1.04/M$0.85/M▼ 18.5%
- Output price$2.08/M$1.70/M▼ 18.5%
- Cache read$0.087/M$0.071/M▼ 18.5%
via merge-gateway
What changed
- structured outputNoYesEnabled
via hyper
What changed
- Input price$0.418/M$0.484/M▲ 15.8%
- Output price$1.59/M$1.85/M▲ 16.6%
- cache write$0.209/M$0.242/M▲ 15.8%
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
- Output limit236K66K−72%
via github-copilot
What changed
Removed from the model catalog
via cortecs
What changed
- Context40K32K−20%
- Output limit40K32K−20%
- Input price$0.099/M$0.089/M▼ 10.1%
- Output price$0.299/M$0.312/M▲ 4.3%
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via kilo
What changed
- nameLing 3.0 Flash FininclusionAI: Ling 3.0 Flash Fin
via crusoe
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via merge-gateway
What changed
- structured outputNoYesEnabled
via openrouter
What changed
First observed in the model catalog
via opencode
What changed
First observed in the model catalog
via venice
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via hyper
What changed
- Input price$0.528/M$0.514/M▼ 2.6%
- Output price$2.79/M$2.75/M▼ 1.1%
- cache write$0.264/M$0.257/M▼ 2.6%
via nano-gpt
What changed
- Context131K33K−75%
- Input limit131K33K−75%
via llmgateway-providers
What changed
- Input price$0.13/M$0.05/M▼ 61.5%
- Output price$0.27/M$0.1/M▼ 63.0%
- Cache read$0.02/M$0.01/M▼ 50.0%
via kimi-for-coding
What changed
- attachmentNoYesEnabled
via scnet-token-plan
What changed
First observed in the model catalog
via github-copilot
What changed
Removed from the model catalog
via fireworks-ai
What changed
First observed in the model catalog
via kimi-for-coding
What changed
- attachmentNoYesEnabled
via kilo
What changed
- Input price$2.50/M$2.55/M▲ 2.0%
- Output price$14/M$12.75/M▼ 8.9%
- Cache read$0.29/M$0.256/M▼ 11.7%
via merge-gateway
What changed
- structured outputNoYesEnabled
via scnet-token-plan
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via llmgateway-providers
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
- Cache read$0.0030/M$0.03/M▲ 900.0%
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via nano-gpt
What changed
- Context33K8K−75%
- Input limit33K8K−75%
via cortecs
What changed
Removed from the model catalog
via github-copilot
What changed
Removed from the model catalog
via openrouter
What changed
First observed in the model catalog
via edenai
What changed
- nameGemini 3.7 FlashGemini 3.8 Flash
- descriptionHigh-efficiency Gemini model for agentic workflows, coding, and multimodal reasoningGoogle's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
- knowledge2026-03—
- release date2026-08-132026-09-02
- +1 more changes
via kilo
What changed
- Output limit393K131K−67%
- Input price$0.05/M$0.05/M▼ 0.0%
- Output price$0.16/M$0.1/M▼ 37.5%
- Cache read$0.013/M$0.0100/M▼ 23.1%
via llmgateway
What changed
First observed in the model catalog
via github-copilot
What changed
Removed from the model catalog
via kilo
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via vercel
What changed
- cache write$0.25/M$0.5/M▲ 100.0%
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.087/M$0.09/M▲ 2.9%
- Output price$0.35/M$0.55/M▲ 57.1%
- Cache read$0.018/M—
via crossmodel
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
- +1 more changes
via crossmodel
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$1.50/M$0.75/M▼ 50.0%
via github-copilot
What changed
- Input price$2/M$4/M▲ 100.0%
- Output price$10/M$20/M▲ 100.0%
- Cache read$0.2/M$0.4/M▲ 100.0%
- cache write$2.50/M$5/M▲ 100.0%
- +1 more changes
via merge-gateway
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via wandb
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context41K262K6.4×
- Output limit33K16K−50%
- Input limit41K262K6.4×
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- nameLing-3.0-flashLing 3.0 Flash
via openrouter
What changed
- Input price$1.12/M$0.579/M▼ 48.0%
- Output price$3.35/M$1.74/M▼ 48.0%
- Cache read$0.037/M$0.019/M▼ 48.0%
via vercel
What changed
- cache write$2.50/M$5/M▲ 100.0%
via nano-gpt
What changed
- Output limit32K64K2×
via openrouter
What changed
- Input price$0.44/M$0.22/M▼ 50.0%
- Output price$1.32/M$0.66/M▼ 50.0%
- Cache read$0.014/M$0.0070/M▼ 50.0%
via vercel
What changed
- cache write$2.50/M$5/M▲ 100.0%
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via edenai
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via crusoe
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via merge-gateway
What changed
- structured outputNoYesEnabled
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- nameLing 3.0 Flash Fin (free)inclusionAI: Ling 3.0 Flash Fin (free)
via llmgateway
What changed
- Input price$0.13/M$0.1/M▼ 23.1%
- Output price$0.4/M$0.25/M▼ 37.5%
- Cache read$0.024/M$0.02/M▼ 16.7%
via merge-gateway
What changed
- structured outputNoYesEnabled
via cortecs
What changed
- Input price$0.25/M$1.01/M▲ 305.6%
- Output price$0.747/M$1.01/M▲ 35.7%
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
- Input price$0.44/M$0.22/M▼ 50.0%
- Output price$1.32/M$0.66/M▼ 50.0%
- Cache read$0.014/M$0.0070/M▼ 50.0%
via vercel
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
- nameLing-3.0-flashinclusionAI: Ling 3.0 Flash
via edenai
What changed
- nameGPT-5.6 SolGPT-6 Astra
- descriptionFrontier GPT-5.6 model for complex professional work, coding, and agentic workflowsGPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
- familygpt-solgpt-astra
- knowledge2026-02-16—
- +7 more changes
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit236K66K−72%
- Input price$0.6/M$0.3/M▼ 50.0%
- Output price$3.60/M$2/M▼ 44.4%
- Cache read$0.12/M$0.03/M▼ 75.0%
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via edenai
What changed
- Input price$0.174/M$0.174/M▲ 0.1%
- Output price$0.755/M$0.755/M▲ 0.1%
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via opencode
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via hyper
What changed
- Input price$0.93/M$0.86/M▼ 7.5%
- Output price$2.88/M$2.78/M▼ 3.3%
- cache write$0.465/M$0.43/M▼ 7.5%
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via crossmodel
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$2.50/M$2.55/M▲ 2.0%
- Output price$14/M$12.75/M▼ 8.9%
- Cache read$0.29/M$0.256/M▼ 11.7%
via fireworks-ai
What changed
- last updated2026-08-262026-09-04
via kilo
What changed
Removed from the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via opencode
What changed
First observed in the model catalog
via nano-gpt
What changed
- Output limit32K64K2×
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via kilo
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via amd
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via merge-gateway
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via edenai
What changed
- Input price$0.465/M$0.465/M▲ 0.1%
- Output price$0.929/M$0.93/M▲ 0.1%
via kilo
What changed
First observed in the model catalog
via llmgateway
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via merge-gateway
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via llmgateway
What changed
- Input price$0.051/M$0.05/M▼ 2.0%
- Output price$0.104/M$0.1/M▼ 3.8%
- Cache read$0.0097/M$0.01/M▲ 3.1%
via merge-gateway
What changed
- structured outputNoYesEnabled
via hyper
What changed
- Input price$1.29/M$1.26/M▼ 2.2%
- Output price$4.22/M$4.13/M▼ 2.1%
- cache write$0.645/M$0.631/M▼ 2.2%
via edenai
What changed
- Input price$0.174/M$0.174/M▲ 0.1%
- Output price$0.697/M$0.697/M▲ 0.1%
via merge-gateway
What changed
- structured outputNoYesEnabled
via cortecs
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via venice
What changed
First observed in the model catalog
via github-copilot
What changed
Removed from the model catalog
via opencode
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.755/M$0.755/M▲ 0.1%
- Output price$0.755/M$0.755/M▲ 0.1%
via llmgateway-providers
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via github-copilot
What changed
Removed from the model catalog
via nano-gpt
What changed
- Context1.05M66K−94%
- Output limit66K33K−50%
- Input limit1.05M66K−94%
via openrouter
What changed
- Output limit393K131K−67%
- Input price$0.05/M$0.05/M▼ 0.0%
- Output price$0.16/M$0.1/M▼ 37.5%
- Cache read$0.013/M$0.0100/M▼ 23.1%
via nano-gpt
What changed
Removed from the model catalog
via crossmodel
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
- Context1.05M1.05M−0%
- Input limit1.05M1.05M−0%
via merge-gateway
What changed
- attachmentYesNoRemoved
- structured outputNoYesEnabled
- modalities.inputtext,imagetext
- Context256K131K−49%
- +1 more changes
via merge-gateway
What changed
- structured outputNoYesEnabled
via merge-gateway
What changed
- structured outputNoYesEnabled
via edenai
What changed
- nameGemini 3.7 FlashGemini 3.8 Flash
- descriptionHigh-efficiency Gemini model for agentic workflows, coding, and multimodal reasoningGoogle's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
- knowledge2026-03—
- release date2026-08-132026-09-02
- +1 more changes
via opencode-go
What changed
First observed in the model catalog
via merge-gateway
What changed
- structured outputNoYesEnabled
via openrouter
What changed
- Input price$1.02/M$1.04/M▲ 1.9%
- Output price$2.04/M$2.08/M▲ 1.9%
- Cache read$0.085/M$0.087/M▲ 1.9%
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via openrouter
What changed
- Input price$0.078/M$0.089/M▲ 13.8%
- Output price$0.156/M$0.177/M▲ 13.8%
- Cache read$0.016/M$0.018/M▲ 13.8%
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via openrouter
What changed
- Output limit4K6K1.3×
- Input price$0.45/M$0.35/M▼ 22.2%
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.66/M$1.12/M▲ 69.0%
- Output price$1.98/M$3.35/M▲ 69.0%
- Cache read$0.022/M$0.037/M▲ 69.0%
via openrouter
What changed
- Output limit183K33K−82%
- Input price$0.6/M$0.625/M▲ 4.2%
- Output price$2.40/M$3.13/M▲ 30.2%
- Cache read$0.12/M$0.188/M▲ 56.3%
via kilo
What changed
- Output limit4K6K1.3×
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via kilo
What changed
- Context203K256K1.3×
- Output limit183K33K−82%
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context32K128K4×
- Output limit33K115K3.5×
- Input limit32K128K4×
via openrouter
What changed
- Output limit228K236K1×
- Cache read$0.025/M$0.03/M▲ 20.0%
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Context262K1.05M4×
- Output limit236K944K4×
- Input price$1.15/M$1.09/M▼ 5.0%
- Output price$3.50/M$3.43/M▼ 1.9%
- +1 more changes
via meta
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit236K944K4×
- Input price$1.15/M$1.09/M▼ 5.0%
- Output price$3.50/M$3.43/M▼ 1.9%
- Cache read$0.1/M$0.179/M▲ 79.4%
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Context200K205K1×
- Output limit205K131K−36%
- Input limit200K205K1×
via openrouter
What changed
- Output limit131K262K2×
- Cache read$0.2/M$0.25/M▲ 25.0%
via nano-gpt
What changed
- Context64K66K1×
- Output limit96K16K−83%
- Input limit64K66K1×
via nano-gpt
What changed
- Output limit1.05M944K−10%
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via empiriolabs
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.66/M$1.12/M▲ 69.0%
- Output price$1.98/M$3.35/M▲ 69.0%
- Cache read$0.022/M$0.037/M▲ 69.0%
via nano-gpt
What changed
- Output limit500K450K−10%
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via llmgateway
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.85/M$0.93/M▲ 9.4%
- Output price$2.77/M$2.88/M▲ 3.7%
- cache write$0.425/M$0.465/M▲ 9.4%
via amd
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context128K131K1×
- Output limit262K33K−88%
- Input limit128K131K1×
via ofox
What changed
First observed in the model catalog
via nano-gpt
What changed
- Output limit1M900K−10%
via empiriolabs
What changed
First observed in the model catalog
via kilo
What changed
- Output limit236K16K−93%
- Input price$0.1/M$0.05/M▼ 50.0%
- Output price$0.9/M$0.7/M▼ 22.2%
- Cache read$0.05/M—
via edenai
What changed
- Input price$0.174/M$0.174/M▲ 0.3%
- Output price$0.695/M$0.697/M▲ 0.3%
via kilo
What changed
- nameMeta: Muse Spark 1.3Muse Spark 1.3
via kilo
What changed
- Context128K164K1.3×
- Output limit16K16K1×
- Input price$0.257/M$0.32/M▲ 24.3%
- Output price$1.03/M$0.89/M▼ 13.5%
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via nano-gpt
What changed
- Context128K164K1.3×
- Output limit164K33K−80%
- Input limit128K164K1.3×
via nano-gpt
What changed
- Context128K131K1×
- Output limit128K118K−8%
- Input limit128K131K1×
via llmgateway
What changed
First observed in the model catalog
via deepinfra
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via openrouter
What changed
- Output limit33K145K4.4×
- Input price$0.25/M$0.55/M▲ 120.0%
- Output price$0.95/M$1.65/M▲ 73.7%
- Cache read$0.13/M$0.55/M▲ 323.1%
via edenai
What changed
- Input price$0.174/M$0.174/M▲ 0.3%
- Output price$0.753/M$0.755/M▲ 0.3%
via nano-gpt
What changed
- descriptionQwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding
- familyqwen3.8-maxqwen
via openrouter
What changed
- Output limit115K16K−86%
- Input price$0.71/M$0.1/M▼ 85.9%
- Output price$0.71/M$0.32/M▼ 54.9%
- Cache read$0.71/M—
via nano-gpt
What changed
- tool callNoYesEnabled
via nano-gpt
What changed
- Context256K262K1×
- Output limit262K236K−10%
- Input limit256K262K1×
via nano-gpt
What changed
- Context127K128K1×
- Output limit128K115K−10%
- Input limit127K128K1×
via nan
What changed
nan began listing this model
via llmgateway-providers
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via ofox
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Context262K1M3.8×
- Output limit131K262K2×
via nano-gpt
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit944K131K−86%
- Input price$0.075/M$0.071/M▼ 5.0%
- Output price$0.25/M$0.237/M▼ 5.0%
- Cache read$0.015/M$0.014/M▼ 5.0%
via nano-gpt
What changed
- Output limit500K450K−10%
via edenai
What changed
- Input price$0.16/M$0.15/M▼ 6.3%
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via tinfoil
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context127K128K1×
- Output limit128K115K−10%
- Input limit127K128K1×
via edenai
What changed
- structured outputNoYesEnabled
via nano-gpt
What changed
- Output limit262K128K−51%
via kilo
What changed
- Context128K131K1×
- Output limit115K16K−86%
via nan
What changed
nan began listing this model
via nano-gpt
What changed
- Context128K131K1×
- Input limit128K131K1×
via vercel
What changed
- familyqwen3.8-maxqwen
- release date2026-09-012026-09-02
- last updated2026-09-012026-09-02
via nano-gpt
What changed
- Context32K33K1×
- Output limit33K26K−20%
- Input limit32K33K1×
via nano-gpt
What changed
- Output limit256K230K−10%
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context64K66K1×
- Output limit96K16K−83%
- Input limit64K66K1×
via nano-gpt
What changed
- Context60K128K2.1×
- Output limit128K115K−10%
- Input limit60K128K2.1×
via nan
What changed
nan began listing this model
via openrouter
What changed
- Output limit131K262K2×
- Cache read$0.26/M$0.14/M▼ 46.2%
via openrouter
What changed
- Input price$0.425/M$0.42/M▼ 1.2%
- Output price$2.55/M$3/M▲ 17.6%
- cache write$0.531/M—
via vercel
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit66K236K3.6×
- Input price$0.39/M$0.55/M▲ 41.0%
- Output price$2.34/M$3.50/M▲ 49.6%
- Cache read—$0.225/M
via edenai
What changed
- Input price$0.463/M$0.465/M▲ 0.3%
- Output price$0.926/M$0.929/M▲ 0.3%
via huggingface
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context32K33K1×
- Output limit33K29K−10%
- Input limit32K33K1×
via nano-gpt
What changed
- Context1M1.05M1×
via llmgateway-providers
What changed
First observed in the model catalog
via edenai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$1.04/M$1.05/M▲ 0.3%
- Output price$1.04/M$1.05/M▲ 0.3%
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Output limit131K262K2×
via nano-gpt
What changed
- Context256K262K1×
- Output limit256K66K−74%
- Input limit256K262K1×
via openrouter
What changed
- Output limit183K33K−82%
- Input price$0.6/M$0.625/M▲ 4.2%
- Output price$2.40/M$3.13/M▲ 30.2%
- Cache read$0.12/M$0.188/M▲ 56.3%
via orcarouter
What changed
- Context1M1.05M1×
- Output limit32K131K4.1×
via nan
What changed
nan began listing this model
via openrouter
What changed
- Output limit29K115K4×
- Input price$0.25/M$0.8/M▲ 220.0%
- Output price$0.75/M$1/M▲ 33.3%
- Cache read—$0.4/M
via nano-gpt
What changed
- Output limit1.05M66K−94%
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context200K205K1×
- Output limit205K131K−36%
- Input limit200K205K1×
via nano-gpt
What changed
- Context256K262K1×
- Output limit256K66K−74%
- Input limit256K262K1×
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Context8K33K4×
- Output limit8K26K3.2×
- Input limit8K33K4×
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Output limit262K236K−10%
via llmgateway
What changed
- Output limit32K131K4.1×
via nan
What changed
nan began listing this model
via vercel
What changed
- Input price$1.40/M$0.7/M▼ 50.0%
- Output price$4.40/M$2.20/M▼ 50.0%
- Cache read$0.14/M$0.13/M▼ 7.1%
via llmgateway
What changed
First observed in the model catalog
via amd
What changed
First observed in the model catalog
via nano-gpt
What changed
- Output limit1.05M944K−10%
via nano-gpt
What changed
- Output limit131K118K−10%
via nano-gpt
What changed
- Output limit131K118K−10%
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Context33K131K4×
- Input limit33K131K4×
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via kilo
What changed
- Input price$0.1/M$0.06/M▼ 40.0%
- Output price$0.15/M$0.25/M▲ 66.7%
- Cache read$0.05/M$0.015/M▼ 70.0%
via meta
What changed
- Output limit32K131K4.1×
via kilo
What changed
- Output limit228K236K1×
- Cache read$0.025/M$0.03/M▲ 20.0%
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.1/M$0.05/M▼ 50.0%
- Output price$0.9/M$0.7/M▼ 22.2%
- Cache read$0.05/M—
via nano-gpt
What changed
- Context256K262K1×
- Output limit262K100K−62%
- Input limit256K262K1×
via llmgateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit16K131K8×
- Input price$0.43/M$0.55/M▲ 27.9%
- Output price$1.75/M$2.20/M▲ 25.7%
- Cache read$0.08/M$0.11/M▲ 37.5%
via kilo
What changed
- Input price$2.55/M$2.50/M▼ 2.0%
- Output price$12.75/M$14/M▲ 9.8%
- Cache read$0.256/M$0.29/M▲ 13.3%
via nano-gpt
What changed
- Output limit131K16K−88%
via meta
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.079/M$0.078/M▼ 1.9%
- Output price$0.159/M$0.156/M▼ 1.9%
- Cache read$0.016/M$0.016/M▼ 1.9%
via nan
What changed
nan began listing this model
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via vercel
What changed
- structured output—YesEnabled
via edenai
What changed
- Input price$0.753/M$0.755/M▲ 0.3%
- Output price$0.753/M$0.755/M▲ 0.3%
via kilo
What changed
- Context41K131K3.2×
- Output limit16K8K−50%
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via nano-gpt
What changed
- Context4K4K1×
- Output limit4K4K−10%
- Input limit4K4K1×
via nano-gpt
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit16K118K7.2×
via nano-gpt
What changed
- Output limit66K8K−88%
via nan
What changed
nan began listing this model
via kilo
What changed
- Context164K161K−2%
- Output limit33K145K4.4×
via nano-gpt
What changed
Removed from the model catalog
via hyper
What changed
- Input price$0.426/M$0.418/M▼ 1.9%
- Output price$1.62/M$1.59/M▼ 2.0%
- cache write$0.213/M$0.209/M▼ 1.9%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Context33K66K2×
- Input limit33K66K2×
via openrouter
What changed
- Input price$2.55/M$2.50/M▼ 2.0%
- Output price$12.75/M$14/M▲ 9.8%
- Cache read$0.256/M$0.29/M▲ 13.3%
via nano-gpt
What changed
- Output limit66K52K−20%
via nano-gpt
What changed
- Output limit66K8K−88%
via openrouter
What changed
- Output limit16K8K−50%
- Input price$0.12/M$0.228/M▲ 89.6%
- Output price$0.24/M$0.91/M▲ 279.2%
via nano-gpt
What changed
- Output limit256K102K−60%
via openrouter
What changed
- Output limit16K118K7.2×
- Output price$1.20/M$1.10/M▼ 8.3%
via tinfoil
What changed
Removed from the model catalog
via kilo
What changed
- Context198K205K1×
- Output limit16K131K8×
- Input price$0.43/M$0.55/M▲ 27.9%
- Output price$1.75/M$2.20/M▲ 25.7%
- +1 more changes
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Context32K128K4×
- Output limit29K115K4×
- Input price$0.25/M$0.8/M▲ 220.0%
- Output price$0.75/M$1/M▲ 33.3%
- +1 more changes
via kilo
What changed
- Output limit66K236K3.6×
via kilo
What changed
- Output limit131K944K7.2×
via nano-gpt
What changed
- Output limit1.05M393K−63%
via nano-gpt
What changed
- descriptionMeta's Muse Spark 1.3 is a frontier multimodal reasoning model for long-horizon coding and agentic workflows, with strong gains in computer use, browsing, professional tool use, codebase understanding, instruction following, and million-token retrieval. It accepts text, images, audio, video, and files, supports tool calling and structured output, and always reasons before answering.Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.
- modalities.inputtext,image,video,audio,pdftext,image,video,pdf,audio
via nano-gpt
What changed
- Context128K131K1×
- Output limit131K16K−88%
- Input limit128K131K1×
via kilo
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context256K262K1×
- Output limit262K236K−10%
- Input limit256K262K1×
via nano-gpt
What changed
- Context256K131K−49%
- Output limit262K118K−55%
- Input limit256K131K−49%
via openrouter
What changed
- Output limit131K944K7.2×
via openrouter
What changed
- Output limit944K131K−86%
- Input price$0.075/M$0.071/M▼ 5.0%
- Output price$0.25/M$0.237/M▼ 5.0%
- Cache read$0.015/M$0.014/M▼ 5.0%
via nano-gpt
What changed
- Cache read$1/M$0.25/M▼ 75.0%
via nano-gpt
What changed
- Output limit131K118K−10%
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Output limit131K102K−22%
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.116/M$0.106/M▼ 8.6%
- Output price$0.38/M$0.368/M▼ 3.2%
- cache write$0.058/M$0.053/M▼ 8.6%
via nano-gpt
What changed
- Output limit500K450K−10%
via openrouter
What changed
- Output limit16K16K1×
- Input price$0.257/M$0.32/M▲ 24.3%
- Output price$1.03/M$0.89/M▼ 13.5%
via openrouter
What changed
- Input price$0.87/M$1.04/M▲ 19.8%
- Output price$1.74/M$2.08/M▲ 19.8%
- Cache read$0.072/M$0.087/M▲ 19.8%
via nano-gpt
What changed
- Context256K262K1×
- Output limit262K236K−10%
- Input limit256K262K1×
via nano-gpt
What changed
- Context127K127K1×
- Output limit128K114K−11%
- Input limit127K127K1×
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via nano-gpt
What changed
- Context262K524K2×
- Input limit262K524K2×
via nebius
What changed
- Context1M1.05M1×
- Output limit384K1.05M2.7×
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via openrouter
What changed
- Input limit—192K
via nebius
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit1M128K−87%
via tencent-tokenhub
What changed
- Output limit64K128K2×
- Input limit—192K
via kilo
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.19/M$0.188/M▼ 1.1%
- Output price$0.63/M$0.7/M▲ 11.1%
- cache write$0.095/M$0.094/M▼ 1.1%
via nebius
What changed
Removed from the model catalog
via deepinfra
What changed
- Output limit64K128K2×
- Input limit—192K
via opencode
What changed
First observed in the model catalog
via llmgateway
What changed
- Output limit64K128K2×
- Input limit—192K
via cortecs
What changed
- Output limit128K16K−88%
via nebius
What changed
First observed in the model catalog
via llmgateway
What changed
- Input price$0.05/M$0.04/M▼ 20.0%
- Output price$0.2/M$0.19/M▼ 5.0%
- Cache read—$0.01/M
via nebius
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit1.05M66K−94%
via edenai
What changed
- Cache read$0.15/M$0.015/M▼ 90.0%
via kilo
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.91/M$0.85/M▼ 6.6%
- Output price$2.93/M$2.77/M▼ 5.5%
- cache write$0.455/M$0.425/M▼ 6.6%
via requesty
What changed
- Input limit—192K
via edenai
What changed
- Input price$0.464/M$0.463/M▼ 0.1%
- Output price$0.927/M$0.926/M▼ 0.1%
via hyper
What changed
- Input price$0.404/M$0.426/M▲ 5.4%
- Output price$1.50/M$1.62/M▲ 8.3%
- cache write$0.202/M$0.213/M▲ 5.4%
via nebius
What changed
- Context256K262K1×
- Input limit256K262K1×
via cortecs
What changed
- Output limit32K8K−74%
via cortecs
What changed
- Output limit198K203K1×
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit1.05M33K−97%
via cortecs
What changed
- Output limit1M128K−87%
via vercel
What changed
- Input limit—192K
via cortecs
What changed
- Output limit1.05M128K−88%
via ovhcloud
What changed
- Input price—$0/M
- Output price—$0/M
via requesty
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit262K256K−2%
via hyper
What changed
- Input price$0.12/M$0.116/M▼ 3.3%
- Output price$0.42/M$0.38/M▼ 9.5%
- cache write$0.06/M$0.058/M▼ 3.3%
via kilo
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit262K262K1×
via nebius
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via kilo
What changed
- Context1.05M1.02M−2%
via hyper
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit400K128K−68%
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit1M128K−87%
via berget
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit1M128K−87%
via opencode-go
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via neon
What changed
- Input price—$1/M
- Output price—$4.05/M
- Cache read—$0.17/M
via cortecs
What changed
- Output limit256K262K1×
via requesty
What changed
- Context131K1.05M8×
via cortecs
What changed
- Input price$1.75/M$1.40/M▼ 20.0%
- Output price$4.50/M$4.40/M▼ 2.2%
- Cache read$0.438/M$0.26/M▼ 40.6%
via cortecs
What changed
- Output limit262K82K−69%
via requesty
What changed
- Output limit384K131K−66%
via meta
What changed
- Context1M1.05M1×
via fireworks-ai
What changed
- Cache read$0.029/M$0.03/M▲ 3.4%
via requesty
What changed
First observed in the model catalog
via berget
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit250K262K1×
via nebius
What changed
First observed in the model catalog
via berget
What changed
Removed from the model catalog
via requesty
What changed
- reasoningNoYesEnabled
via cortecs
What changed
- Output limit1.05M128K−88%
via nebius
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit196K197K1×
via nebius
What changed
Removed from the model catalog
via nebius
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.1%
- Output price$0.753/M$0.753/M▼ 0.1%
via cortecs
What changed
- Output limit300K66K−78%
via tencent-token-plan
What changed
- Output limit64K128K2×
- Input limit—192K
via openrouter
What changed
- Output limit944K236K−75%
- Input price$1.17/M$1.15/M▼ 1.7%
- Output price$3.96/M$3.50/M▼ 11.6%
- Cache read$0.234/M$0.1/M▼ 57.3%
via cortecs
What changed
- Output limit128K16K−88%
via cortecs
What changed
- Output limit131K262K2×
via hyper
What changed
- Input price$0.55/M$0.528/M▼ 4.0%
- Output price$2.88/M$2.79/M▼ 3.5%
- cache write$0.275/M$0.264/M▼ 4.0%
via requesty
What changed
- Context1.05M1M−5%
- Output limit131K262K2×
- Input price$0.075/M$0.2/M▲ 166.7%
- Output price$0.25/M$0.5/M▲ 100.0%
- +1 more changes
via openrouter
What changed
- Output limit131K393K3×
- Input price$0.05/M$0.05/M▲ 0.0%
- Output price$0.1/M$0.16/M▲ 60.1%
- Cache read$0.0100/M$0.013/M▲ 30.1%
via edenai
What changed
First observed in the model catalog
via gitlab
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit200K64K−68%
via berget
What changed
Removed from the model catalog
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Output limit64K128K2×
- Input limit—192K
via nebius
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit32K4K−87%
via nebius
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Context131K66K−50%
via openrouter
What changed
- Input price$1.03/M$1.02/M▼ 0.3%
- Output price$2.05/M$2.05/M▼ 0.3%
- Cache read$0.086/M$0.085/M▼ 0.3%
via nano-gpt
What changed
Removed from the model catalog
via nebius
What changed
Removed from the model catalog
via berget
What changed
- last updated2026-08-292026-09-01
- Context328K524K1.6×
via cortecs
What changed
- Output limit1.05M66K−94%
via nebius
What changed
First observed in the model catalog
via edenai
What changed
- Cache read$0.15/M$0.015/M▼ 90.0%
via openrouter
What changed
- Input price$0.072/M$0.07/M▼ 2.1%
- Output price$0.143/M$0.14/M▼ 2.1%
- Cache read$0.014/M$0.014/M▼ 2.1%
via nebius
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit1M66K−93%
via kilo
What changed
- Context1.05M262K−75%
- Output limit944K236K−75%
- Input price$1.17/M$1.15/M▼ 1.7%
- Output price$3.96/M$3.50/M▼ 11.6%
- +1 more changes
via cortecs
What changed
- Output limit131K110K−16%
via nano-gpt
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit262K131K−50%
- Cache read$0.25/M$0.2/M▼ 20.0%
via openai
What changed
- status—deprecated
via cortecs
What changed
- Output limit1.05M66K−94%
via edenai
What changed
- Cache read$0.07/M$0.0070/M▼ 90.0%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit197K196K−0%
via openrouter
What changed
- Output limit16K33K2×
- Input price$0.5/M$0.625/M▲ 25.0%
- Output price$2.20/M$3.13/M▲ 42.0%
- Cache read$0.1/M$0.188/M▲ 87.5%
via cortecs
What changed
- Output limit1.05M128K−88%
via merge-gateway
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit64K1.05M16.4×
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via opencode
What changed
Removed from the model catalog
via huggingface
What changed
- Output limit64K128K2×
- Input limit—192K
via nebius
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit1.05M128K−88%
via requesty
What changed
- reasoningNoYesEnabled
- Output limit384K131K−66%
- Input price$0.076/M$0.14/M▲ 84.2%
- Output price$0.153/M$0.28/M▲ 83.0%
- +1 more changes
via nebius
What changed
- Context33K41K1.3×
- Input limit33K41K1.3×
via nebius
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via vercel
What changed
First observed in the model catalog
via llmgateway
What changed
First observed in the model catalog
via fireworks-ai
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via nebius
What changed
First observed in the model catalog
via cortecs
What changed
- Output limit32K8K−74%
via nebius
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via neon
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$12/M▼ 20.0%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write—$2.50/M
- +1 more changes
via cortecs
What changed
- Output limit66K16K−75%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via opencode
What changed
- Input limit—192K
via venice
What changed
First observed in the model catalog
via neon
What changed
- Input price$1/M$0.2/M▼ 80.0%
- Output price$6/M$1.20/M▼ 80.0%
- Cache read$0.1/M$0.02/M▼ 80.0%
- cache write—$0.25/M
- +1 more changes
via cortecs
What changed
- Output limit200K64K−68%
via openrouter
What changed
First observed in the model catalog
via hyper
What changed
- Input price$1.33/M$1.29/M▼ 3.2%
- Output price$4.31/M$4.22/M▼ 2.1%
- cache write$0.666/M$0.645/M▼ 3.2%
via nano-gpt
What changed
- Input price$0.375/M$0.75/M▲ 100.0%
- Output price$1.88/M$3.75/M▲ 100.0%
- Cache read$0.037/M$0.075/M▲ 100.0%
- cache write$0.021/M$0.075/M▲ 260.0%
via cortecs
What changed
- Output limit400K196K−51%
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit1.05M66K−94%
via openrouter
What changed
First observed in the model catalog
via nebius
What changed
Removed from the model catalog
via nebius
What changed
Removed from the model catalog
via edenai
What changed
- nameClaude Fable 5Claude Fable 5.1
- descriptionClaude model for creative writing, analysis, and controlled agent workflowsClaude model for demanding reasoning and long-horizon agentic work
- knowledge2026-01-312026-06
- release date2026-06-092026-09-01
- +2 more changes
via edenai
What changed
- Input price$0.753/M$0.753/M▼ 0.1%
- Output price$0.753/M$0.753/M▼ 0.1%
via edenai
What changed
- Input price$1.04/M$1.04/M▼ 0.1%
- Output price$1.04/M$1.04/M▼ 0.1%
via crossmodel
What changed
- Input limit—192K
via cortecs
What changed
- Output limit1.05M64K−94%
via kilo
What changed
- Context161K164K1×
- Output limit145K33K−77%
via azure
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
- +1 more changes
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Output limit262K131K−50%
via kilo
What changed
- Output limit131K393K3×
- Input price$0.05/M$0.05/M▲ 0.0%
- Output price$0.1/M$0.16/M▲ 60.1%
- Cache read$0.0100/M$0.013/M▲ 30.1%
via nano-gpt
What changed
Removed from the model catalog
via nebius
What changed
- Context128K131K1×
via nebius
What changed
- Context128K262K2×
- Input limit120K262K2.2×
via nebius
What changed
Removed from the model catalog
via kilo
What changed
- Context262K256K−2%
- Output limit16K33K2×
via opencode-go
What changed
First observed in the model catalog
via openrouter
What changed
- familygeminigemini-flash
- modalities.inputtext,image,video,pdf,audiotext,image,video,audio,pdf
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- +3 more changes
via llmgateway-providers
What changed
First observed in the model catalog
via opencode-go
What changed
- Output limit64K128K2×
- Input limit—192K
via cortecs
What changed
- Output limit1.05M66K−94%
via opencode
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via crossmodel
What changed
First observed in the model catalog
via requesty
What changed
First observed in the model catalog
via kenari
What changed
- Output limit64K128K2×
- Input limit—192K
via cortecs
What changed
- Output limit128K131K1×
via opencode
What changed
- statusdeprecated—
via cortecs
What changed
- Output limit1.05M33K−97%
via nebius
What changed
- Context128K131K1×
via cortecs
What changed
- Output limit200K64K−68%
via requesty
What changed
First observed in the model catalog
via ovhcloud
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via nebius
What changed
Removed from the model catalog
via nebius
What changed
Removed from the model catalog
via requesty
What changed
- structured outputYesNoRemoved
via berget
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via nano-gpt
What changed
- descriptionGoogle's fast multimodal model for agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Its capabilities, limits, reasoning behavior, and provisional pricing currently mirror Gemini 3.7 Flash.Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
- familygeminigemini-flash
- Input price$0.375/M$0.75/M▲ 100.0%
- Output price$1.88/M$3.75/M▲ 100.0%
- +2 more changes
via kilo
What changed
- Input limit—192K
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit5K10K2×
via edenai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via opencode
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via llmgateway-providers
What changed
- Input limit—192K
via requesty
What changed
- reasoningNoYesEnabled
via cortecs
What changed
- Output limit131K128K−2%
via openrouter
What changed
- Output limit262K131K−50%
- Input price$1.19/M$0.966/M▼ 18.8%
- Output price$3.74/M$3.04/M▼ 18.8%
- Cache read$0.221/M$0.193/M▼ 12.6%
via cortecs
What changed
- Output limit128K10K−92%
via kilo
What changed
First observed in the model catalog
via requesty
What changed
First observed in the model catalog
via orcarouter
What changed
- Output limit64K128K2×
- Input limit—192K
via cortecs
What changed
- Output limit1M128K−87%
via kilo
What changed
First observed in the model catalog
via azure
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.375/M$0.75/M▲ 100.0%
- Output price$1.88/M$3.75/M▲ 100.0%
- Cache read$0.037/M$0.075/M▲ 100.0%
- cache write$0.021/M$0.075/M▲ 260.0%
via google
What changed
First observed in the model catalog
via berget
What changed
Removed from the model catalog
via requesty
What changed
- reasoningNoYesEnabled
via neon
What changed
- Input price$0.072/M$0.15/M▲ 108.3%
- Output price$0.28/M$0.6/M▲ 114.3%
via nebius
What changed
Removed from the model catalog
via openai
What changed
- status—deprecated
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via kenari
What changed
- Output limit64K128K2×
- Input limit—192K
via cortecs
What changed
- Output limit1.05M1M−5%
via ovhcloud
What changed
- Input price—$0/M
- Output price—$0/M
via cortecs
What changed
- Output limit300K10K−97%
via cortecs
What changed
- Output limit200K65K−68%
via cortecs
What changed
- Output limit1.05M33K−97%
via google-vertex
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Cache read$0.07/M$0.0070/M▼ 90.0%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.1%
- Output price$0.695/M$0.695/M▼ 0.1%
via requesty
What changed
- reasoningNoYesEnabled
- structured outputNoYesEnabled
- Context1M1.05M1×
- Output limit384K131K−66%
- +3 more changes
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit256K262K1×
via openrouter
What changed
- Output limit145K33K−77%
- Input price$0.55/M$0.25/M▼ 54.5%
- Output price$1.65/M$0.95/M▼ 42.4%
- Cache read$0.55/M$0.13/M▼ 76.4%
via nano-gpt
What changed
Removed from the model catalog
via llmgateway
What changed
- Input price$1.30/M$1.20/M▼ 7.7%
- Cache read$0.25/M$0.2/M▼ 20.0%
via opencode-go
What changed
- statusdeprecated—
via cortecs
What changed
- Output limit1M128K−87%
via cortecs
What changed
- Output limit400K128K−68%
via neon
What changed
- Input price$0.05/M$0.07/M▲ 40.0%
- Output price$0.2/M$0.3/M▲ 50.0%
via openrouter
What changed
- Input price$0.66/M$1.12/M▲ 69.0%
- Output price$1.98/M$3.35/M▲ 69.0%
- Cache read$0.022/M$0.037/M▲ 69.0%
via berget
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via orcarouter
What changed
- Output limit64K128K2×
- Input limit—192K
via nano-gpt
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit400K128K−68%
via cortecs
What changed
- Output limit400K128K−68%
via llmgateway-providers
What changed
- Input limit—192K
via nano-gpt
What changed
Removed from the model catalog
via nebius
What changed
Removed from the model catalog
via cortecs
What changed
- Output limit128K4K−97%
via nebius
What changed
- Context432K1.05M2.4×
- Output limit432K1.05M2.4×
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- nameGoogle: Gemini 3.8 FlashGemini 3.8 Flash
- familygeminigemini-flash
- modalities.inputtext,image,video,pdf,audiotext,image,video,audio,pdf
- Input price$1.50/M$0.75/M▼ 50.0%
- +4 more changes
via cortecs
What changed
- Output limit262K256K−2%
via neon
What changed
- cache write—$6.25/M
- tiers[object Object][object Object]
via cortecs
What changed
- Output limit1.05M66K−94%
via cortecs
What changed
- Output limit262K33K−87%
via jalapeno
What changed
- Output limit64K128K2×
- Input limit—192K
via kilo
What changed
Removed from the model catalog
via llmgateway
What changed
First observed in the model catalog
via kilo
What changed
- nameNex AGI: Nex-N2-MiniNex AGI: Nex-N2-Mini (retires Sep 4)
via iteracompute
What changed
First observed in the model catalog
via anthropic
What changed
First observed in the model catalog
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.1%
- Output price$0.696/M$0.695/M▼ 0.1%
via amazon-bedrock
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Output limit393K131K−67%
- Input price$0.05/M$0.05/M▼ 0.0%
- Output price$0.16/M$0.1/M▼ 37.5%
- Cache read$0.013/M$0.0100/M▼ 23.1%
via kilo
What changed
- modalities.inputtext,imagetext,image,video
- Output limit1.05M944K−10%
- Cache read$0/M—
- reasoning$0/M—
via digitalocean
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$3/M$2.50/M▼ 16.7%
via kilo
What changed
- Context1M262K−74%
- Output limit262K131K−50%
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.5/M$0.4/M▼ 20.0%
via kilo
What changed
- Context1.02M1.05M1×
- Output limit384K393K1×
via kilo
What changed
- Output limit197K177K−10%
- Cache read$0/M—
- reasoning$0/M—
via kilo
What changed
- nameAnthropic: Claude Fable 5.1 ($$$$)Claude Fable 5.1
- knowledge—2026-06
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via merge-gateway
What changed
First observed in the model catalog
via kilo
What changed
- nameNex AGI: Nex-N2-ProNex AGI: Nex-N2-Pro (retires Sep 4)
via cortecs
What changed
- Context262K256K−2%
- Input price$0.446/M$0.478/M▲ 7.2%
- Output price$2.23/M$2.39/M▲ 7.4%
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
- Context1.05M1.02M−2%
via opencode
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via edenai
What changed
- Input price$0.464/M$0.464/M▼ 0.1%
- Output price$0.928/M$0.927/M▼ 0.1%
via hyper
What changed
- Input price$0.424/M$0.404/M▼ 4.7%
- Output price$1.61/M$1.50/M▼ 7.2%
- cache write$0.212/M$0.202/M▼ 4.7%
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via requesty
What changed
- tiers[object Object]
via vercel
What changed
First observed in the model catalog
via vercel
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit384K393K1×
- Input price$0.87/M$1.60/M▲ 83.9%
- Output price$1.74/M$3.20/M▲ 83.9%
- Cache read$0.072/M$0.135/M▲ 86.2%
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via edenai
What changed
- Input price$0.174/M$0.174/M▼ 0.1%
- Output price$0.754/M$0.753/M▼ 0.1%
via llmgateway-providers
What changed
First observed in the model catalog
via chutes
What changed
- Input price$0.35/M$0.32/M▼ 8.6%
- Output price$2.75/M$2.50/M▼ 9.1%
- Cache read$0.035/M$0.032/M▼ 8.6%
via kilo
What changed
- structured outputNoYesEnabled
via edenai
What changed
- Input price$0.754/M$0.753/M▼ 0.1%
- Output price$0.754/M$0.753/M▼ 0.1%
via cortecs
What changed
Removed from the model catalog
via vercel
What changed
- temperatureYesNoRemoved
- knowledge—2026-06
- release date2026-08-312026-09-01
- last updated2026-08-312026-09-01
via hyper
What changed
- Input price$0.544/M$0.55/M▲ 1.1%
- Output price$2.85/M$2.88/M▲ 1.1%
- cache write$0.272/M$0.275/M▲ 1.1%
via kilo
What changed
- Input price$0.45/M$0.35/M▼ 22.2%
via opencode
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.66/M$1.12/M▲ 69.0%
- Output price$1.98/M$3.35/M▲ 69.0%
- Cache read$0.022/M$0.037/M▲ 69.0%
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via edenai
What changed
- Input price$1.04/M$1.04/M▼ 0.1%
- Output price$1.04/M$1.04/M▼ 0.1%
via edenai
What changed
- Input price$0.08/M$0.03/M▼ 62.5%
- Output price$0.18/M$0.1/M▼ 44.4%
via nano-gpt
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via amazon-bedrock
What changed
First observed in the model catalog
via venice
What changed
- Input price$10/M$12/M▲ 20.0%
- Output price$50/M$60/M▲ 20.0%
- Cache read$0.25/M$0.3/M▲ 20.0%
- cache write$12.50/M$15/M▲ 20.0%
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
- knowledge—2026-06
via requesty
What changed
- tiers[object Object]
via kilo
What changed
- Input price$3/M$2.50/M▼ 16.7%
via openrouter
What changed
- Output limit8K16K2×
- Input price$0.11/M$0.1/M▼ 9.1%
- Output price$0.34/M$0.3/M▼ 11.8%
- Cache read$0.055/M—
via nano-gpt
What changed
- descriptionClaude Fable 5.1 improves on Fable 5 across agentic coding, long-running workflows, front-end and visual code generation, finance, analysis, and knowledge work, with more concise plans and summaries. Anthropic retains prompts and outputs for 30 days; Zero Data Retention is not available.Claude model for demanding reasoning and long-horizon agentic work
- temperatureYesNoRemoved
- knowledge—2026-06
via kilo
What changed
- Cache read$1/M$0.25/M▼ 75.0%
via requesty
What changed
- tiers[object Object]
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.087/M$0.09/M▲ 2.9%
- Output price$0.35/M$0.55/M▲ 57.1%
- Cache read$0.018/M—
via fireworks-ai
What changed
- Input price$0.14/M$0.22/M▲ 57.1%
- Output price$0.28/M$0.66/M▲ 135.7%
- Cache read$0.028/M$0.0070/M▼ 75.0%
via openrouter
What changed
- structured outputNoYesEnabled
via openrouter
What changed
First observed in the model catalog
via kilo
What changed
Removed from the model catalog
via requesty
What changed
First observed in the model catalog
via vancine
What changed
Removed from the model catalog
via vercel
What changed
- Context1M1.05M1×
- Output limit384K1.05M2.7×
via openrouter
What changed
- Output limit393K131K−67%
- Input price$0.05/M$0.05/M▼ 0.0%
- Output price$0.16/M$0.1/M▼ 37.5%
- Cache read$0.013/M$0.0100/M▼ 23.1%
via iteracompute
What changed
- modalities.inputtext,image,videotext,image
- Context262K328K1.3×
- Output limit33K66K2×
- Input limit—262K
- +3 more changes
via venice
What changed
First observed in the model catalog
via kilo
What changed
Removed from the model catalog
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via requesty
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via azure
What changed
First observed in the model catalog
via google-vertex
What changed
First observed in the model catalog
via azure-cognitive-services
What changed
First observed in the model catalog
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
- Context131K328K2.5×
- Output limit8K16K2×
via kilo
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via kilo
What changed
- Output limit236K16K−93%
via merge-gateway
What changed
- tool callNoYesEnabled
- structured outputNoYesEnabled
- modalities.inputtext,imagetext,image,pdf
via hyper
What changed
- Input price$0.188/M$0.19/M▲ 1.1%
- Output price$0.7/M$0.63/M▼ 10.0%
- cache write$0.094/M$0.095/M▲ 1.1%
via openrouter
What changed
- Cache read$1/M$0.25/M▼ 75.0%
via venice
What changed
- Input price$2/M$3/M▲ 50.0%
- Output price$10/M$15/M▲ 50.0%
- Cache read$0.2/M$0.3/M▲ 50.0%
- cache write$2.50/M$3.75/M▲ 50.0%
via requesty
What changed
- tiers[object Object][object Object]
via openrouter
What changed
- Input price$0.079/M$0.081/M▲ 1.9%
- Output price$0.159/M$0.162/M▲ 1.9%
- Cache read$0.016/M$0.016/M▲ 1.9%
via edenai
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +2 more changes
via openrouter
What changed
- Input price$0.5/M$0.4/M▼ 20.0%
via vancine
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit4K7K2×
- Input price$0.06/M$0.4/M▲ 566.7%
- Output price$0.06/M$0.6/M▲ 900.0%
via requesty
What changed
- tiers[object Object]
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- Context4K8K2×
- Output limit4K7K2×
via google-vertex-anthropic
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via cortecs
What changed
- Context197K196K−0%
via llmgateway-providers
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image,video,pdf
via llmgateway
What changed
Removed from the model catalog
via requesty
What changed
- reasoningYesNoRemoved
- Output limit131K384K2.9×
- Input price$0.14/M$0.076/M▼ 45.7%
- Output price$0.28/M$0.153/M▼ 45.4%
- +1 more changes
via aihubmix
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via kilo
What changed
- Context1.05M128K−88%
- Output limit16K115K7×
via requesty
What changed
- Input price$1.20/M$0.8/M▼ 33.3%
- Output price$4.20/M$2.55/M▼ 39.3%
- Cache read$0.26/M$0.16/M▼ 38.5%
via openrouter
What changed
- Output limit131K944K7.2×
- Input price$1.19/M$1.17/M▼ 1.5%
- Output price$4.18/M$3.96/M▼ 5.3%
- Cache read$0.247/M$0.234/M▼ 5.3%
via llmgateway-providers
What changed
Removed from the model catalog
via hyper
What changed
- Input price$0.9/M$0.91/M▲ 1.1%
- Output price$2.80/M$2.93/M▲ 4.6%
- cache write$0.45/M$0.455/M▲ 1.1%
via openrouter
What changed
- structured outputYesNoRemoved
via aihubmix
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.3/M$0.15/M▼ 50.0%
- Output price$2.50/M$1.25/M▼ 50.0%
- Cache read$0.03/M$0.015/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +1 more changes
via kilo
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$9/M$4.50/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +1 more changes
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via kilo
What changed
- Input price$0.25/M$0.125/M▼ 50.0%
- Output price$1.50/M$0.75/M▼ 50.0%
- Cache read$0.025/M$0.013/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +1 more changes
via kilo
What changed
- Input price$0.5/M$0.25/M▼ 50.0%
- Output price$3/M$1.50/M▼ 50.0%
- Cache read$0.05/M$0.025/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +1 more changes
via openrouter
What changed
Removed from the model catalog
via opencode
What changed
- attachmentYesNoRemoved
via hyper
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway
What changed
- Input price$0.1/M$0.36/M▲ 260.0%
- Output price$0.3/M$0.87/M▲ 190.0%
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway
What changed
Removed from the model catalog
via kilo
What changed
- Context1.02M1.05M1×
- Output limit384K393K1×
via llmgateway
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.4/M$0.257/M▼ 35.6%
- Output price$1.30/M$1.03/M▼ 20.9%
via hyper
What changed
- Input price$0.11/M$0.12/M▲ 9.1%
- Output price$0.408/M$0.42/M▲ 2.9%
- cache write$0.055/M$0.06/M▲ 9.1%
via edenai
What changed
- Input price$0.175/M$0.174/M▼ 0.4%
- Output price$0.757/M$0.754/M▼ 0.4%
via vancine
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via llmgateway
What changed
Removed from the model catalog
via empiriolabs
What changed
Removed from the model catalog
via venice
What changed
- open weightsNoYesEnabled
via llmgateway
What changed
Removed from the model catalog
via google
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit80K236K2.9×
via kilo
What changed
- Context256K262K1×
- Output limit80K144K1.8×
via llmgateway-providers
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image,video,pdf
via kilo
What changed
- Input price$0.75/M$0.375/M▼ 50.0%
- Output price$3.75/M$1.88/M▼ 50.0%
- Cache read$0.075/M$0.037/M▼ 50.0%
- cache write$0.042/M$0.021/M▼ 50.0%
- +1 more changes
via kilo
What changed
- structured outputYesNoRemoved
- Input price$0.22/M$0.25/M▲ 13.6%
- Output price$0.85/M$0.8/M▼ 5.9%
via deepinfra
What changed
- Cache read$0.24/M$0.12/M▼ 50.0%
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- structured outputYesNoRemoved
- Context1M1.05M1×
- Output limit128K1.05M8.2×
- Input price$1.40/M$1.20/M▼ 14.3%
- +1 more changes
via hyper
What changed
- Input price$0.404/M$0.424/M▲ 5.0%
- Output price$1.50/M$1.61/M▲ 7.8%
- cache write$0.202/M$0.212/M▲ 5.0%
via llmgateway
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via llmgateway-providers
What changed
Removed from the model catalog
via crossmodel
What changed
- Input price$1.50/M$1.88/M▲ 25.0%
- Output price$4.50/M$5.63/M▲ 25.0%
- Cache read$0.3/M$0.375/M▲ 25.0%
- cache write$1.88/M$2.35/M▲ 25.0%
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit16K236K14.4×
- Input price$0.09/M$0.1/M▲ 11.1%
- Cache read—$0.07/M
via vercel
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via kilo
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +1 more changes
via coralbricks
What changed
Removed from the model catalog
via hyper
What changed
- Input price$0.284/M$0.274/M▼ 3.5%
- Output price$0.934/M$0.899/M▼ 3.7%
- cache write$0.142/M$0.137/M▼ 3.5%
via hyper
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via llmgateway-providers
What changed
Removed from the model catalog
via aihubmix
What changed
First observed in the model catalog
via llmgateway
What changed
Removed from the model catalog
via kilo
What changed
- Context1M262K−74%
- Output limit262K131K−50%
via requesty
What changed
- structured outputNoYesEnabled
- Context1M131K−87%
- Output limit128K131K1×
- Input price$0.15/M$0.075/M▼ 50.0%
- +2 more changes
via edenai
What changed
- Input price$0.466/M$0.464/M▼ 0.4%
- Output price$0.931/M$0.928/M▼ 0.4%
via vancine
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.079/M$0.09/M▲ 14.2%
- Output price$0.157/M$0.18/M▲ 14.2%
- Cache read$0.016/M$0.018/M▲ 14.2%
via merge-gateway
What changed
- cache write—$0.25/M
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via crossmodel
What changed
- Output limit1.05M66K−94%
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
- Input price$1.05/M$1.04/M▼ 0.4%
- Output price$1.05/M$1.04/M▼ 0.4%
via openrouter
What changed
First observed in the model catalog
via venice
What changed
- open weightsNoYesEnabled
via llmgateway-providers
What changed
Removed from the model catalog
via merge-gateway
What changed
- cache write—$5/M
via crossmodel
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.3/M$0.15/M▼ 50.0%
- Output price$2.50/M$1.25/M▼ 50.0%
- Cache read$0.03/M$0.015/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
via openrouter
What changed
- Input price$0.6/M$0.45/M▼ 25.0%
- Output price$3/M$2.25/M▼ 25.0%
- Cache read$0.1/M$0.07/M▼ 30.0%
via crossmodel
What changed
- Input price$0.288/M$0.32/M▲ 11.1%
- Output price$1.13/M$1.25/M▲ 11.1%
- Cache read$0.029/M$0.032/M▲ 11.1%
- cache write$0.36/M$0.4/M▲ 11.1%
- +1 more changes
via llmgateway-providers
What changed
Removed from the model catalog
via vancine
What changed
First observed in the model catalog
via abliteration-ai
What changed
First observed in the model catalog
via venice
What changed
- open weightsNoYesEnabled
via requesty
What changed
- structured outputYesNoRemoved
- Input price$1.75/M$1.20/M▼ 31.4%
- Output price$4.50/M$4.20/M▼ 6.7%
- Cache read$0.44/M$0.26/M▼ 40.9%
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway
What changed
- Input price$0.13/M$0.135/M▲ 3.8%
via llmgateway-providers
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.25/M$0.125/M▼ 50.0%
- Output price$1.50/M$0.75/M▼ 50.0%
- Cache read$0.025/M$0.013/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
- +1 more changes
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Input price$2/M$1/M▼ 50.0%
- Output price$12/M$6/M▼ 50.0%
- Cache read$0.2/M$0.1/M▼ 50.0%
- cache write$0.375/M$0.188/M▼ 50.0%
- +1 more changes
via merge-gateway
What changed
- Input price$2/M$3/M▲ 50.0%
- Output price$10/M$15/M▲ 50.0%
via merge-gateway
What changed
- Input price$1.50/M$2/M▲ 33.3%
- Output price$4.50/M$6/M▲ 33.3%
- Cache read$0.375/M$0.5/M▲ 33.3%
via kilo
What changed
Removed from the model catalog
via groq
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via vancine
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.03/M$0.05/M▲ 66.7%
via openrouter
What changed
- Output limit80K144K1.8×
via llmgateway
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit262K131K−50%
- Cache read$0.25/M$0.2/M▼ 20.0%
via llmgateway
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$1.50/M$1.65/M▲ 10.0%
- Output price$9/M$9.90/M▲ 10.0%
- Cache read$0.15/M$0.165/M▲ 10.0%
- reasoning$9/M$9.90/M▲ 10.0%
via aihubmix
What changed
First observed in the model catalog
via coralbricks
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via edenai
What changed
- Input price$0.175/M$0.174/M▼ 0.4%
- Output price$0.699/M$0.696/M▼ 0.4%
via kilo
What changed
- Input price$2/M$1/M▼ 50.0%
- Output price$12/M$6/M▼ 50.0%
- Cache read$0.2/M$0.1/M▼ 50.0%
- cache write$0.375/M$0.188/M▼ 50.0%
- +1 more changes
via aihubmix
What changed
First observed in the model catalog
via crossmodel
What changed
- Output limit262K131K−50%
via venice
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via merge-gateway
What changed
- cache write—$2.50/M
via hyper
What changed
- Input price$1.31/M$1.33/M▲ 1.4%
- Output price$4.27/M$4.31/M▲ 1.0%
- cache write$0.657/M$0.666/M▲ 1.4%
via llmgateway-providers
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.42/M$1/M▲ 138.1%
- Output price$1.32/M$3.20/M▲ 142.4%
- Cache read$0.078/M$0.2/M▲ 156.4%
via moonshotai-cn
What changed
- attachmentNoYesEnabled
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via hyper
What changed
- Input price$0.178/M$0.188/M▲ 5.6%
- Output price$0.68/M$0.7/M▲ 2.9%
- cache write$0.089/M$0.094/M▲ 5.6%
via nano-gpt
What changed
- Input price$0.06/M$0.05/M▼ 16.7%
- Output price$0.3/M$0.25/M▼ 16.7%
- Cache read$0.03/M$0.025/M▼ 16.7%
via kilo
What changed
- Context262K1.05M4×
- Output limit131K944K7.2×
- Input price$1.19/M$1.17/M▼ 1.5%
- Output price$4.18/M$3.96/M▼ 5.3%
- +1 more changes
via llmgateway
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.66/M$1.32/M▲ 100.0%
- Output price$1.98/M$3.96/M▲ 100.0%
- Cache read$0.022/M$0.044/M▲ 100.0%
via kilo
What changed
- Output limit16K236K14.4×
via llmtech
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via kilo
What changed
- Context256K262K1×
- Output limit80K236K2.9×
via cortecs
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit384K393K1×
- Input price$0.417/M$1.60/M▲ 283.5%
- Output price$0.835/M$3.20/M▲ 283.5%
- Cache read$0.035/M$0.135/M▲ 288.3%
via kilo
What changed
- Input price$0.27/M$0.25/M▼ 7.4%
- Output price$1.12/M$1/M▼ 10.7%
- Cache read$0.135/M—
via edenai
What changed
- Input price$0.757/M$0.754/M▼ 0.4%
- Output price$0.757/M$0.754/M▼ 0.4%
via openrouter
What changed
- Output limit16K115K7×
- Output price$0.8/M$0.696/M▼ 13.0%
via kilo
What changed
Removed from the model catalog
via aihubmix
What changed
First observed in the model catalog
via merge-gateway
What changed
First observed in the model catalog
via llmgateway
What changed
Removed from the model catalog
via aihubmix
What changed
First observed in the model catalog
via kilo
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.03/M$0.05/M▲ 66.7%
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via merge-gateway
What changed
- Input price$3/M$2.90/M▼ 3.3%
- Output price$15/M$14/M▼ 6.7%
via moonshotai
What changed
- attachmentNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Output limit182K128K−30%
- Input price$1.26/M$0.966/M▼ 23.3%
- Output price$3.96/M$3.04/M▼ 23.3%
- Cache read$0.234/M$0.179/M▼ 23.3%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- release date2026-08-212026-07-29
via cortecs
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.2/M$0.15/M▼ 25.0%
- Output price$1.40/M$0.7/M▼ 50.0%
via nano-gpt
What changed
Removed from the model catalog
via above
What changed
above began listing this model
via nano-gpt
What changed
Removed from the model catalog
via sensenova
What changed
sensenova began listing this model
via nano-gpt
What changed
- release date2026-08-242026-07-29
via nano-gpt
What changed
- release date2026-08-282026-07-29
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via edenai
What changed
- Cache read—$0.15/M
via above
What changed
above began listing this model
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via cline-pass
What changed
- open weightsNoYesEnabled
via edenai
What changed
- open weightsNoYesEnabled
via klokintegration
What changed
klokintegration began listing this model
via berget
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- release date2026-08-262026-07-29
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.2/M$0.15/M▼ 25.0%
- Output price$1.40/M$0.7/M▼ 50.0%
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.165/M$0.019/M▼ 88.5%
- Output price$0.165/M$0.03/M▼ 81.8%
- Cache read$0.017/M—
via nano-gpt
What changed
- Input price$0.05/M$0.1/M▲ 100.0%
- Output price$0.1/M$0.2/M▲ 100.0%
- Cache read$0.025/M$0.05/M▲ 100.0%
via above
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit943K33K−97%
- Cache read$0.03/M$0.025/M▼ 16.7%
via nano-gpt
What changed
Removed from the model catalog
via trustedrouter
What changed
Removed from the model catalog
via edenai
What changed
- Cache read—$0.15/M
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.1/M$0.09/M▼ 10.0%
- Cache read$0.07/M—
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via zai
What changed
- open weightsNoYesEnabled
via aiand
What changed
- open weightsNoYesEnabled
via trustedrouter
What changed
First observed in the model catalog
via zhipuai
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via crof
What changed
- open weightsNoYesEnabled
via above
What changed
above began listing this model
via nano-gpt
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- release date2026-08-212026-07-29
via cloudflare-workers-ai
What changed
- Context1.31M1.05M−20%
- Output limit1.31M1.05M−20%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via bothub
What changed
bothub began listing this model
via zhipuai-coding-plan
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.125/M$0.35/M▲ 180.0%
- Output price$0.5/M$1.40/M▲ 180.0%
- Cache read$0.063/M$0.175/M▲ 180.0%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via togetherai
What changed
- open weightsNoYesEnabled
via requesty
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via trustedrouter
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via above
What changed
above began listing this model
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via nano-gpt
What changed
- structured outputYesNoRemoved
via tokenrouter
What changed
- open weightsNoYesEnabled
via openrouter
What changed
- Input price$0.529/M$0.508/M▼ 3.8%
- Output price$1.06/M$1.02/M▼ 3.8%
- Cache read$0.044/M$0.042/M▼ 3.9%
via nano-gpt
What changed
- release date2026-08-232026-07-29
via nano-gpt
What changed
Removed from the model catalog
via fireworks-ai
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via deepinfra
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$13.50/M$10/M▼ 25.9%
- Cache read$0.25/M$0.2/M▼ 20.0%
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via openrouter
What changed
- Output limit66K147K2.3×
- Input price$0.269/M$0.26/M▼ 3.3%
- Output price$0.4/M$0.38/M▼ 5.0%
- Cache read$0.135/M$0.13/M▼ 3.3%
via neuralwatt
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.44/M$0.4/M▼ 9.1%
- Output price$2.20/M$2/M▼ 9.1%
- Cache read$0.044/M$0.04/M▼ 9.1%
via venice
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via above
What changed
above began listing this model
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit66K147K2.3×
via synthetic
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via nano-gpt
What changed
- open weightsNoYesEnabled
via trustedrouter
What changed
Removed from the model catalog
via openrouter
What changed
- Output price$0.14/M$0.16/M▲ 14.3%
via nano-gpt
What changed
Removed from the model catalog
via above
What changed
above began listing this model
via nano-gpt
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Output limit236K16K−93%
via bothub
What changed
bothub began listing this model
via nano-gpt
What changed
First observed in the model catalog
via cloudflare-workers-ai
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via berget
What changed
First observed in the model catalog
via openrouter
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.2/M$0.25/M▲ 25.0%
- Output price$1.40/M$1.50/M▲ 7.1%
- Cache read$0.04/M$0.125/M▲ 212.5%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Context203K200K−1%
- Output limit182K128K−30%
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via trustedrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via baseten
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
First observed in the model catalog
via aiand
What changed
First observed in the model catalog
via digitalocean
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- structured outputYesNoRemoved
via kilo
What changed
- Input price$0.11/M$0.075/M▼ 31.8%
- Output price$0.33/M$0.2/M▼ 39.4%
- Cache read$0.011/M—
via nano-gpt
What changed
- release date2026-08-252026-07-29
via nano-gpt
What changed
Removed from the model catalog
via fireworks-ai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.18/M$0.25/M▲ 38.9%
- Output price$0.5/M$1.50/M▲ 200.0%
- Cache read$0.075/M$0.125/M▲ 66.7%
via nano-gpt
What changed
- release date2026-08-242026-07-29
via nano-gpt
What changed
Removed from the model catalog
via trustedrouter
What changed
Removed from the model catalog
via trustedrouter
What changed
First observed in the model catalog
via tokengo
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via vercel
What changed
- open weightsNoYesEnabled
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via trustedrouter
What changed
First observed in the model catalog
via kilo
What changed
- Output limit943K33K−97%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via llmgateway
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via trustedrouter
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit236K80K−66%
- Input price$0.22/M$0.25/M▲ 13.6%
- Output price$0.85/M$0.8/M▼ 5.9%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- release date2026-08-232026-07-29
via nano-gpt
What changed
- Output price$0.5/M$0.95/M▲ 90.0%
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via kenari
What changed
- open weightsNoYesEnabled
via trustedrouter
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via above
What changed
above began listing this model
via nano-gpt
What changed
- release date2026-08-242026-07-29
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Output price$1.20/M$1.10/M▼ 8.3%
via nano-gpt
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$13.50/M$10/M▼ 25.9%
- Cache read$0.25/M$0.2/M▼ 20.0%
via vancine
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.05/M$0.1/M▲ 100.0%
- Output price$0.1/M$0.2/M▲ 100.0%
- Cache read$0.025/M$0.05/M▲ 100.0%
via nano-gpt
What changed
First observed in the model catalog
via trustedrouter
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit16K115K7×
- Output price$0.8/M$0.696/M▼ 13.0%
via nano-gpt
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.44/M$0.4/M▼ 9.1%
- Output price$2.20/M$2/M▼ 9.1%
- Cache read$0.044/M$0.04/M▼ 9.1%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via trustedrouter
What changed
First observed in the model catalog
via klokintegration
What changed
klokintegration began listing this model
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via crossmodel
What changed
- open weightsNoYesEnabled
via kilo
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.18/M$0.25/M▲ 38.9%
- Output price$0.5/M$1.50/M▲ 200.0%
- Cache read$0.075/M$0.125/M▲ 66.7%
via nano-gpt
What changed
- Output price$0.5/M$0.95/M▲ 90.0%
via orcarouter
What changed
- open weightsNoYesEnabled
via sensenova
What changed
sensenova began listing this model
via zhipuai-coding-plan
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.18/M$0.25/M▲ 38.9%
- Output price$0.5/M$1.50/M▲ 200.0%
- Cache read$0.075/M$0.125/M▲ 66.7%
via nano-gpt
What changed
Removed from the model catalog
via sensenova
What changed
sensenova began listing this model
via zai-coding-plan
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via kilo
What changed
- Output price$0.1/M$0.14/M▲ 40.0%
- Cache read$0.0070/M$0.01/M▲ 42.9%
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- release date2026-08-202026-07-29
via friendli
What changed
- open weightsNoYesEnabled
via volcengine-coding-plan
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via zai-coding-plan
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.18/M$0.25/M▲ 38.9%
- Output price$0.5/M$1.50/M▲ 200.0%
- Cache read$0.075/M$0.125/M▲ 66.7%
via nano-gpt
What changed
First observed in the model catalog
via empiriolabs
What changed
- open weightsNoYesEnabled
via trustedrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via kilo
What changed
- Context1.05M128K−88%
- Output limit16K115K7×
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via edenai
What changed
- Cache read—$0.03/M
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit236K80K−66%
via klokintegration
What changed
klokintegration began listing this model
via nano-gpt
What changed
First observed in the model catalog
via aiand
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via openrouter
What changed
- Input price$0.081/M$0.081/M▼ 0.7%
- Output price$0.163/M$0.162/M▼ 0.7%
- Cache read$0.016/M$0.016/M▼ 0.7%
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via nano-gpt
What changed
Removed from the model catalog
via requesty
What changed
- open weightsNoYesEnabled
via edenai
What changed
- Cache read—$0.015/M
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via above
What changed
above began listing this model
via nano-gpt
What changed
- Context16K33K2×
- Input limit16K33K2×
via tokenrouter
What changed
tokenrouter began listing this model
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via nano-gpt
What changed
Removed from the model catalog
via opencode-go
What changed
- nameHy3 (8x usage)Hy3
- Input price$0.018/M$0.14/M▲ 700.0%
- Output price$0.072/M$0.58/M▲ 700.0%
- Cache read$0.0044/M$0.035/M▲ 700.0%
via nano-gpt
What changed
Removed from the model catalog
via merge-gateway
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via ollama-cloud
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via requesty
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via vivgrid
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- attachmentNoYesEnabled
- release date2026-08-272026-07-29
- modalities.inputtexttext,image
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- release date2026-08-202026-07-29
via opencode-go
What changed
- open weightsNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via ofox
What changed
- open weightsNoYesEnabled
via kilo
What changed
- Output limit262K472K1.8×
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- release date2026-08-242026-07-29
via nano-gpt
What changed
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.33/M$0.45/M▲ 36.4%
- Cache read$0.04/M$0.05/M▲ 25.0%
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.44/M$0.4/M▼ 9.1%
- Output price$2.20/M$2/M▼ 9.1%
- Cache read$0.044/M$0.04/M▼ 9.1%
via nano-gpt
What changed
Removed from the model catalog
via trustedrouter
What changed
Removed from the model catalog
via opencode
What changed
- status—deprecated
via nano-gpt
What changed
- Input price$0.08/M$0.12/M▲ 50.0%
- Output price$0.33/M$0.38/M▲ 15.2%
- Cache read$0.04/M$0.06/M▲ 50.0%
via requesty
What changed
Removed from the model catalog
via nano-gpt
What changed
- release date2026-08-242026-07-29
via nano-gpt
What changed
- release date2026-08-262026-07-29
via nano-gpt
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
Removed from the model catalog
via friendli
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$0.14/M$0.22/M▲ 57.1%
- Output price$0.28/M$0.66/M▲ 135.7%
- Cache read$0.028/M$0.0070/M▼ 75.0%
via kilo
What changed
- Output limit66K147K2.3×
via openrouter
What changed
- Input price$0.087/M$0.087/M▼ 0.3%
- Output price$0.174/M$0.173/M▼ 0.3%
- Cache read$0.017/M$0.017/M▼ 0.3%
via openrouter
What changed
- Output limit131K944K7.2×
- Input price$1.25/M$1.20/M▼ 4.0%
- Output price$4.40/M$4/M▼ 9.1%
- Cache read$0.26/M$0.24/M▼ 7.7%
via kilo
What changed
- Context262K1.05M4×
- Output limit131K944K7.2×
- Input price$1.25/M$1.20/M▼ 4.0%
- Output price$4.40/M$4/M▼ 9.1%
- +1 more changes
via llmgateway-providers
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via edenai
What changed
- Context1.05M1.02M−2%
via kilo
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.083/M$0.14/M▲ 69.7%
- Output price$0.33/M$0.58/M▲ 75.8%
- Cache read$0.021/M$0.035/M▲ 69.7%
via kilo
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.742/M$0.731/M▼ 1.4%
- Output price$1.48/M$1.46/M▼ 1.4%
- Cache read$0.062/M$0.061/M▼ 1.4%
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via llmgateway-providers
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.03/M$0.02/M▼ 33.3%
- Output price$0.13/M$0.1/M▼ 23.1%
via openrouter
What changed
- Output price$0.1/M$0.14/M▲ 40.0%
- Cache read$0.0070/M$0.01/M▲ 42.9%
via crof
What changed
First observed in the model catalog
via vercel
What changed
- Input price$0.55/M$0.5/M▼ 9.1%
- Output price$3.30/M$3/M▼ 9.1%
- Cache read$0.11/M$0.1/M▼ 9.1%
- cache write—$0.625/M
via openrouter
What changed
Removed from the model catalog
via edenai
What changed
- tool callNoYesEnabled
via openrouter
What changed
- Output limit66K147K2.3×
- Input price$0.269/M$0.26/M▼ 3.3%
- Output price$0.4/M$0.38/M▼ 5.0%
- Cache read$0.135/M$0.13/M▼ 3.3%
via togetherai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via edenai
What changed
- tool callNoYesEnabled
via edenai
What changed
- tool callNoYesEnabled
via digitalocean
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit131K262K2×
via llmgateway
What changed
- Input price$1.40/M$1.30/M▼ 7.1%
- Output price$4.40/M$4/M▼ 9.1%
- Cache read$0.26/M$0.25/M▼ 3.8%
via kilo
What changed
- Output limit131K262K2×
via kilo
What changed
- Input price$0.08/M$0.07/M▼ 12.5%
- Cache read$0.01/M$0.1/M▲ 900.0%
via deepinfra
What changed
- Input price$1.40/M$1.20/M▼ 14.3%
- Output price$4.40/M$4/M▼ 9.1%
- Cache read$0.26/M$0.24/M▼ 7.7%
via openrouter
What changed
- Input price$0.05/M$0.045/M▼ 10.0%
- Output price$0.1/M$0.09/M▼ 10.0%
- Cache read$0.01/M$0.0090/M▼ 10.0%
via openrouter
What changed
- Output limit262K472K1.8×
- Input price$0.95/M$1/M▲ 5.3%
- Cache read$0.16/M$0.17/M▲ 6.3%
via openrouter
What changed
Removed from the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via openreason
What changed
openreason began listing this model
via neuralwatt
What changed
First observed in the model catalog
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via alibaba
What changed
First observed in the model catalog
via ofox
What changed
- structured output—YesEnabled
via kilo
What changed
- Context1.05M1M−5%
- Output limit131K262K2×
via nano-gpt
What changed
First observed in the model catalog
via nano-gpt
What changed
- tool callNoYesEnabled
via kenari
What changed
First observed in the model catalog
via vancine
What changed
vancine began listing this model
via requesty
What changed
- Input price$13.50/M$15/M▲ 11.1%
- Output price$67.50/M$75/M▲ 11.1%
- Cache read$1.35/M$1.50/M▲ 11.1%
- cache write$16.88/M$18.75/M▲ 11.1%
via kenari
What changed
First observed in the model catalog
via digitalocean
What changed
- Input price$0.068/M$0.14/M▲ 106.2%
- Output price$0.168/M$0.28/M▲ 66.7%
- Cache read$0.017/M$0.028/M▲ 66.7%
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.225/M$0.25/M▲ 11.1%
- Output price$1.80/M$2/M▲ 11.1%
- Cache read$0.045/M$0.05/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$22.50/M$25/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$5.63/M$6.25/M▲ 11.1%
via orcarouter
What changed
Removed from the model catalog
via requesty
What changed
- Input price$9.90/M$11/M▲ 11.1%
- Output price$49.50/M$55/M▲ 11.1%
- Cache read$0.99/M$1.10/M▲ 11.1%
- cache write$12.38/M$13.75/M▲ 11.1%
via volcengine
What changed
- Output limit32K131K4.1×
via vancine
What changed
vancine began listing this model
via openrouter
What changed
- Context1.05M1.31M1.3×
via orcarouter
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$10.80/M$12/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
- cache write$4.05/M$4.50/M▲ 11.1%
- +1 more changes
via requesty
What changed
- Input price$0.396/M$0.44/M▲ 11.1%
- Output price$1.98/M$2.20/M▲ 11.1%
- Cache read$0.396/M$0.44/M▲ 11.1%
via inceptron
What changed
- Input price$0.54/M$0.53/M▼ 1.9%
- Cache read$0.15/M$0.17/M▲ 13.3%
via vancine
What changed
vancine began listing this model
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.98/M$2.20/M▲ 11.1%
- Output price$9.90/M$11/M▲ 11.1%
- Cache read$0.198/M$0.22/M▲ 11.1%
- cache write$2.48/M$2.75/M▲ 11.1%
via openrouter
What changed
- Context1.05M1.31M1.3×
- Output limit944K131K−86%
- Input price$1.40/M$1.25/M▼ 10.7%
via orcarouter
What changed
First observed in the model catalog
via crof
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.98/M$2.20/M▲ 11.1%
- Output price$7.92/M$8.80/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
via orcarouter
What changed
- Output limit262K33K−88%
- cache write—$0/M
via neuralwatt
What changed
- structured output—NoRemoved
- Output limit200K32K−84%
- Input price$0.725/M$0.943/M▲ 30.0%
- Output price$2.25/M$2.92/M▲ 30.0%
- +1 more changes
via nano-gpt
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.396/M$0.44/M▲ 11.1%
- Output price$1.98/M$2.20/M▲ 11.1%
- Cache read$0.396/M$0.44/M▲ 11.1%
via orcarouter
What changed
- Input price$0.14/M$0.147/M▲ 5.0%
- Output price$0.28/M$0.295/M▲ 5.4%
- Cache read$0.028/M$0.02/M▼ 28.6%
via vercel
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via digitalocean
What changed
- Input price$0.4/M$0.8/M▲ 100.0%
- Output price$1.50/M$3/M▲ 100.0%
- Cache read$0.08/M$0.16/M▲ 100.0%
via fireworks-ai
What changed
Removed from the model catalog
via requesty
What changed
- Input price$0.6/M$0.75/M▲ 25.0%
- Output price$3/M$3.75/M▲ 25.0%
- Cache read$0.06/M$0.075/M▲ 25.0%
via opencode-go
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via openreason
What changed
openreason began listing this model
via nano-gpt
What changed
- Input price$1/M$2/M▲ 100.0%
- Output price$6/M$12/M▲ 100.0%
- Cache read$0.1/M$0.2/M▲ 100.0%
- cache write$1.25/M$2.50/M▲ 100.0%
via tokengo
What changed
tokengo began listing this model
via togetherai
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.45/M$0.5/M▲ 11.1%
- Output price$2.70/M$3/M▲ 11.1%
- Cache read$0.045/M$0.05/M▲ 11.1%
- cache write$0.563/M$0.625/M▲ 11.1%
via volcengine
What changed
- structured output—YesEnabled
via requesty
What changed
- Input price$2.25/M$2.50/M▲ 11.1%
- Output price$13.50/M$15/M▲ 11.1%
- Cache read$0.225/M$0.25/M▲ 11.1%
- tiers[object Object][object Object]
via orcarouter
What changed
Removed from the model catalog
via cortecs
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.09/M$0.08/M▼ 11.1%
- Output price$0.34/M$0.35/M▲ 2.9%
- Cache read$0.05/M$0.01/M▼ 80.0%
via orcarouter
What changed
- Output limit128K100K−22%
via nano-gpt
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.45/M$0.5/M▲ 11.1%
- Output price$2.70/M$3/M▲ 11.1%
- Cache read$0.09/M$0.1/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit33K943K28.8×
- Cache read$0.025/M$0.03/M▲ 20.0%
via requesty
What changed
- Input price$0.855/M$0.95/M▲ 11.1%
- Output price$3.60/M$4/M▲ 11.1%
- Cache read$0.171/M$0.19/M▲ 11.1%
via wandb
What changed
Removed from the model catalog
via requesty
What changed
- Input price$1.49/M$1.65/M▲ 11.1%
- Output price$7.42/M$8.25/M▲ 11.1%
- Cache read$1.49/M$1.65/M▲ 11.1%
via vercel
What changed
- Output limit13K1M78.1×
- Cache read$0.26/M$0.14/M▼ 46.2%
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via kilo
What changed
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- Cache read$0.033/M$0.035/M▲ 6.1%
via openrouter
What changed
- Context32K33K1×
- Output limit26K26K1×
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via orcarouter
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via tencent-token-plan
What changed
First observed in the model catalog
via orcarouter
What changed
- Input price$4/M$2/M▼ 50.0%
- Output price$18/M$12/M▼ 33.3%
via llmgateway
What changed
- Input price$3/M$2.83/M▼ 5.7%
- Output price$15/M$14.13/M▼ 5.8%
- Cache read$0.3/M$0.28/M▼ 6.7%
via kilo
What changed
- Output limit236K82K−65%
via ofox
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.9/M$1/M▲ 11.1%
- Output price$4.50/M$5/M▲ 11.1%
- Cache read$0.09/M$0.1/M▲ 11.1%
- cache write$1.13/M$1.25/M▲ 11.1%
via volcengine
What changed
- reasoningNoYesEnabled
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.19/M$1.32/M▲ 11.1%
- Output price$3.56/M$3.96/M▲ 11.1%
- Cache read$0.04/M$0.044/M▲ 11.1%
via requesty
What changed
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$2.25/M$2.50/M▲ 11.1%
via ofox
What changed
- structured output—YesEnabled
via nano-gpt
What changed
First observed in the model catalog
via orcarouter
What changed
- Output limit262K33K−88%
via volcengine
What changed
- Output limit32K131K4.1×
via requesty
What changed
- Input price$1.68/M$1.87/M▲ 11.1%
- Output price$4.21/M$4.68/M▲ 11.1%
- Cache read$0.337/M$0.374/M▲ 11.1%
via neuralwatt
What changed
- structured output—NoRemoved
- Output limit200K32K−84%
- Cache read$0.362/M$0.145/M▼ 60.0%
via requesty
What changed
- Input price$1.49/M$1.65/M▲ 11.1%
- Output price$7.42/M$8.25/M▲ 11.1%
- Cache read$1.49/M$1.65/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via requesty
What changed
- Input price$0.675/M$0.75/M▲ 11.1%
- Output price$4.05/M$4.50/M▲ 11.1%
- Cache read$0.068/M$0.075/M▲ 11.1%
via neuralwatt
What changed
Removed from the model catalog
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$22.50/M$25/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$5.63/M$6.25/M▲ 11.1%
via openrouter
What changed
- Output limit147K66K−56%
- Input price$0.26/M$0.269/M▲ 3.5%
- Output price$0.38/M$0.4/M▲ 5.3%
- Cache read$0.13/M$0.135/M▲ 3.5%
via requesty
What changed
- Input price$0.148/M$0.165/M▲ 11.1%
- Output price$0.594/M$0.66/M▲ 11.1%
- Cache read$0.074/M$0.083/M▲ 11.1%
via orcarouter
What changed
- Cache read—$0.02/M
via orcarouter
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via requesty
What changed
- Input price$0.247/M$0.275/M▲ 11.1%
- Output price$1.98/M$2.20/M▲ 11.1%
- Cache read$0.025/M$0.028/M▲ 11.1%
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- Input price$2.97/M$3.30/M▲ 11.1%
- Output price$14.85/M$16.50/M▲ 11.1%
- Cache read$0.27/M$0.3/M▲ 11.1%
- cache write$3.71/M$4.13/M▲ 11.1%
- +1 more changes
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via requesty
What changed
- Input price$27/M$30/M▲ 11.1%
- Output price$162/M$180/M▲ 11.1%
- Cache read$27/M$30/M▲ 11.1%
via github-copilot
What changed
- cache write—$2.50/M
- tiers[object Object][object Object]
via openrouter
What changed
- Context262K131K−50%
via requesty
What changed
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$1.08/M$1.20/M▲ 11.1%
via runinfra
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.126/M$0.14/M▲ 11.1%
- Output price$0.252/M$0.28/M▲ 11.1%
- Cache read$0.063/M$0.07/M▲ 11.1%
via vivgrid
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.99/M$1.10/M▲ 11.1%
- Output price$4.95/M$5.50/M▲ 11.1%
- Cache read$0.099/M$0.11/M▲ 11.1%
- cache write$1.24/M$1.38/M▲ 11.1%
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$5.40/M$6/M▲ 11.1%
- Cache read$0.225/M$0.25/M▲ 11.1%
- cache write$2.25/M$2.50/M▲ 11.1%
via openrouter
What changed
- Input price$0.05/M$0.06/M▲ 20.0%
- Output price$0.1/M$0.12/M▲ 20.0%
- Cache read$0.01/M$0.012/M▲ 20.0%
via modal
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.375/M$0.75/M▲ 100.0%
- Output price$1.88/M$3.75/M▲ 100.0%
- Cache read$0.037/M$0.075/M▲ 100.0%
- cache write$0.021/M$0.042/M▲ 100.0%
- +1 more changes
via kilo
What changed
- structured outputNoYesEnabled
via requesty
What changed
- Input price$2.70/M$3/M▲ 11.1%
- Output price$13.50/M$15/M▲ 11.1%
- Cache read$0.27/M$0.3/M▲ 11.1%
- cache write$3.38/M$3.75/M▲ 11.1%
- +1 more changes
via edenai
What changed
First observed in the model catalog
via alibaba-token-plan-cn
What changed
First observed in the model catalog
via orcarouter
What changed
- Input price$0.435/M$0.147/M▼ 66.2%
- Output price$0.87/M$0.295/M▼ 66.1%
via fireworks-ai
What changed
Removed from the model catalog
via requesty
What changed
- Input price$1.26/M$1.40/M▲ 11.1%
- Output price$3.96/M$4.40/M▲ 11.1%
- Cache read$1.26/M$1.40/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$1/M$2/M▲ 100.0%
- Output price$6/M$12/M▲ 100.0%
- Cache read$0.1/M$0.2/M▲ 100.0%
- cache write$1.25/M$2.50/M▲ 100.0%
via opencode-go
What changed
First observed in the model catalog
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via vancine
What changed
vancine began listing this model
via requesty
What changed
- Input price$1.24/M$1.38/M▲ 11.1%
- Output price$9.90/M$11/M▲ 11.1%
- Cache read$0.124/M$0.138/M▲ 11.1%
via requesty
What changed
- Input price$0.225/M$0.25/M▲ 11.1%
- Output price$1.35/M$1.50/M▲ 11.1%
- Cache read$0.022/M$0.025/M▲ 11.1%
- cache write$0.075/M$0.083/M▲ 11.1%
- +1 more changes
via orcarouter
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via vercel
What changed
- Context1M512K−49%
- Output limit1M512K−49%
via digitalocean
What changed
- Input price$0.055/M$0.1/M▲ 81.8%
- Output price$0.385/M$0.7/M▲ 81.8%
via requesty
What changed
- Input price$1.35/M$1.50/M▲ 11.1%
- Output price$6.30/M$7/M▲ 11.1%
- Cache read$0.135/M$0.15/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.89/M$2.10/M▲ 11.1%
- Output price$5.94/M$6.60/M▲ 11.1%
- Cache read$0.189/M$0.21/M▲ 11.1%
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via kenari
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.0%
- Output price$0.699/M$0.699/M▼ 0.0%
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via llmgateway
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.198/M$0.22/M▲ 11.1%
- Output price$1.19/M$1.32/M▲ 11.1%
- Cache read$0.02/M$0.022/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via github-copilot
What changed
- cache write—$0.25/M
- tiers[object Object][object Object]
via kilo
What changed
- Input price$0.075/M$0.15/M▲ 100.0%
- Output price$0.25/M$0.5/M▲ 100.0%
- Cache read$0.015/M$0.03/M▲ 100.0%
via vancine
What changed
vancine began listing this model
via neuralwatt
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit131K236K1.8×
- Input price$0.08/M$0.1/M▲ 25.0%
- Output price$0.2/M$0.25/M▲ 25.0%
- Cache read$0.04/M$0.05/M▲ 25.0%
via kenari
What changed
First observed in the model catalog
via openrouter
What changed
- structured outputNoYesEnabled
via ofox
What changed
- structured output—YesEnabled
via runinfra
What changed
- Cache read—$0.01/M
via nano-gpt
What changed
- Input price$0.1/M$0.2/M▲ 100.0%
- Output price$0.6/M$1.20/M▲ 100.0%
- Cache read$0.01/M$0.02/M▲ 100.0%
- cache write$0.125/M$0.25/M▲ 100.0%
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via llmgateway-providers
What changed
Removed from the model catalog
via vercel
What changed
- Context256K262K1×
- Output limit128K262K2×
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- +1 more changes
via neuralwatt
What changed
- structured output—YesEnabled
- Cache read$0.072/M$0.029/M▼ 60.0%
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$5.40/M$6/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$1.80/M$2/M▲ 11.1%
- +1 more changes
via requesty
What changed
- Input price$0.392/M$0.435/M▲ 11.1%
- Output price$0.783/M$0.87/M▲ 11.1%
- Cache read$0.0032/M$0.0036/M▲ 11.1%
via crof
What changed
Removed from the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via neuralwatt
What changed
- reasoningNoYesEnabled
- structured output—NoRemoved
- Cache read$0.362/M$0.145/M▼ 60.0%
via tokengo
What changed
tokengo began listing this model
via requesty
What changed
- Input price$1.19/M$1.32/M▲ 11.1%
- Output price$3.56/M$3.96/M▲ 11.1%
- Cache read$0.04/M$0.044/M▲ 11.1%
via openrouter
What changed
- Context256K262K1×
via neuralwatt
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.45/M$0.5/M▲ 11.1%
- Output price$1.80/M$2/M▲ 11.1%
- tiers[object Object][object Object]
via kilo
What changed
- Output limit33K943K28.8×
via requesty
What changed
- Input price$2.02/M$2.25/M▲ 11.1%
- Output price$10.13/M$11.25/M▲ 11.1%
- Cache read$0.203/M$0.225/M▲ 11.1%
via nano-gpt
What changed
First observed in the model catalog
via ofox
What changed
- structured output—YesEnabled
via requesty
What changed
- Input price$0.09/M$0.1/M▲ 11.1%
- Output price$0.27/M$0.3/M▲ 11.1%
via deepinfra
What changed
First observed in the model catalog
via amd
What changed
First observed in the model catalog
via neuralwatt
What changed
- reasoningNoYesEnabled
- structured output—NoRemoved
- Output limit200K32K−84%
- Input price$0.725/M$0.943/M▲ 30.0%
- +2 more changes
via openrouter
What changed
- Input price$3/M$2.55/M▼ 15.0%
- Output price$15/M$12.75/M▼ 15.0%
- Cache read$0.3/M$0.256/M▼ 14.7%
via requesty
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.24/M$0.3/M▲ 25.0%
- Output price$0.96/M$1.20/M▲ 25.0%
- Cache read$0.048/M$0.06/M▲ 25.0%
via requesty
What changed
- Input price$1.35/M$1.50/M▲ 11.1%
- Output price$8.10/M$9/M▲ 11.1%
- Cache read$0.135/M$0.15/M▲ 11.1%
- cache write$1.42/M$1.58/M▲ 11.1%
via requesty
What changed
- Input price$0.66/M$0.825/M▲ 25.0%
- Output price$3.30/M$4.13/M▲ 25.0%
- Cache read$0.066/M$0.083/M▲ 25.0%
via requesty
What changed
- Input price$0.396/M$0.44/M▲ 11.1%
- Output price$1.98/M$2.20/M▲ 11.1%
- Cache read$0.396/M$0.44/M▲ 11.1%
via requesty
What changed
- Input price$9/M$10/M▲ 11.1%
- Output price$45/M$50/M▲ 11.1%
- Cache read$0.9/M$1/M▲ 11.1%
- cache write$11.25/M$12.50/M▲ 11.1%
via volcengine
What changed
- structured output—YesEnabled
via tokengo
What changed
tokengo began listing this model
via orcarouter
What changed
Removed from the model catalog
via requesty
What changed
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$2.25/M$2.50/M▲ 11.1%
- Cache read$0.027/M$0.03/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via vancine
What changed
vancine began listing this model
via neuralwatt
What changed
- reasoningNoYesEnabled
- structured output—NoRemoved
- Output limit200K32K−84%
- Cache read$0.362/M$0.145/M▼ 60.0%
via fireworks-ai
What changed
Removed from the model catalog
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.855/M$0.95/M▲ 11.1%
- Output price$3.60/M$4/M▲ 11.1%
- Cache read$0.144/M$0.16/M▲ 11.1%
via requesty
What changed
- Input price$0.09/M$0.1/M▲ 11.1%
- Output price$0.36/M$0.4/M▲ 11.1%
- Cache read$0.018/M$0.02/M▲ 11.1%
via neuralwatt
What changed
- structured output—NoRemoved
- Cache read$0.362/M$0.145/M▼ 60.0%
via orcarouter
What changed
- Output limit128K100K−22%
via volcengine
What changed
- structured output—YesEnabled
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.54/M$0.6/M▲ 11.1%
- Output price$2.16/M$2.40/M▲ 11.1%
- Cache read$0.054/M$0.06/M▲ 11.1%
- cache write$1.08/M$1.20/M▲ 11.1%
via regolo-ai
What changed
First observed in the model catalog
via llmgateway
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via edenai
What changed
First observed in the model catalog
via huggingface
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via requesty
What changed
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$1.08/M$1.20/M▲ 11.1%
- Cache read$0.054/M$0.06/M▲ 11.1%
- cache write$1.08/M$1.20/M▲ 11.1%
via orcarouter
What changed
- Input price$0.56/M$0.442/M▼ 21.1%
- Output price$1.12/M$0.884/M▼ 21.1%
- Cache read$0.0036/M$0.06/M▲ 1555.2%
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$24.75/M$27.50/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
- cache write$6.19/M$6.88/M▲ 11.1%
via openrouter
What changed
- Output limit944K384K−59%
- Input price$1.12/M$1.32/M▲ 17.6%
- Output price$3.37/M$3.96/M▲ 17.6%
- Cache read$0.037/M$0.044/M▲ 17.6%
via orcarouter
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via orcarouter
What changed
First observed in the model catalog
via edenai
What changed
- tool callNoYesEnabled
- structured outputNoYesEnabled
via requesty
What changed
- Input price$0.126/M$0.14/M▲ 11.1%
- Output price$0.252/M$0.28/M▲ 11.1%
- Cache read$0.0025/M$0.0028/M▲ 11.1%
via orcarouter
What changed
- Cache read$0.2/M$0.26/M▲ 30.0%
via llmgateway
What changed
- Input price$0.95/M$0.89/M▼ 6.3%
- Output price$4/M$3.71/M▼ 7.3%
- Cache read$0.19/M$0.18/M▼ 5.3%
via requesty
What changed
- Input price$0.18/M$0.2/M▲ 11.1%
- Output price$1.03/M$1.15/M▲ 11.1%
- Cache read$0.036/M$0.04/M▲ 11.1%
via crof
What changed
- Input price$0.12/M$0.08/M▼ 33.3%
- Output price$0.21/M$0.1/M▼ 52.4%
via neuralwatt
What changed
Removed from the model catalog
via openrouter
What changed
- structured outputNoYesEnabled
- Output limit131K944K7.2×
via neuralwatt
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.375/M$0.75/M▲ 100.0%
- Output price$1.88/M$3.75/M▲ 100.0%
- Cache read$0.037/M$0.075/M▲ 100.0%
- cache write$0.021/M$0.042/M▲ 100.0%
- +1 more changes
via edenai
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via volcengine
What changed
- structured output—YesEnabled
via orcarouter
What changed
- Input price$5/M$2.50/M▼ 50.0%
- Output price$22.50/M$15/M▼ 33.3%
via digitalocean
What changed
- Input price$0.7/M$1.40/M▲ 100.0%
- Output price$2.20/M$4.40/M▲ 100.0%
- Cache read$0.105/M$0.21/M▲ 100.0%
via vercel
What changed
- Cache read$0.19/M$0.16/M▼ 15.8%
via requesty
What changed
- Input price$1.24/M$1.38/M▲ 11.1%
- Output price$9.90/M$11/M▲ 11.1%
- Cache read$0.124/M$0.138/M▲ 11.1%
via requesty
What changed
- Input price$0.18/M$0.2/M▲ 11.1%
- Output price$1.08/M$1.20/M▲ 11.1%
- Cache read$0.018/M$0.02/M▲ 11.1%
- tiers[object Object][object Object]
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via edenai
What changed
- Input price$0.466/M$0.466/M▼ 0.0%
- Output price$0.932/M$0.931/M▼ 0.0%
via llmgateway-providers
What changed
First observed in the model catalog
via neuralwatt
What changed
- Context262K262K−0%
- Output limit262K262K−0%
- Input price$0.475/M$0.618/M▲ 30.0%
- Output price$2/M$2.60/M▲ 30.0%
- +1 more changes
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$22.50/M$25/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$5.63/M$6.25/M▲ 11.1%
via neuralwatt
What changed
First observed in the model catalog
via tokengo
What changed
tokengo began listing this model
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.297/M$0.33/M▲ 11.1%
- Output price$2.48/M$2.75/M▲ 11.1%
- Cache read$0.03/M$0.033/M▲ 11.1%
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$29.70/M$33/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
via requesty
What changed
- Input price$2.02/M$2.25/M▲ 11.1%
- Output price$10.13/M$11.25/M▲ 11.1%
- Cache read$0.203/M$0.225/M▲ 11.1%
via nano-gpt
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit131K262K2×
via cloudflare-workers-ai
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.09/M$0.1/M▲ 11.1%
- Output price$0.45/M$0.5/M▲ 11.1%
via requesty
What changed
- Input price$0.36/M$0.4/M▲ 11.1%
- Output price$2.70/M$3/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via edenai
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via neuralwatt
What changed
- structured output—NoRemoved
- Input price$0.725/M$0.943/M▲ 30.0%
- Output price$2.25/M$2.92/M▲ 30.0%
- Cache read$0.181/M$0.094/M▼ 48.0%
via requesty
What changed
- Input price$0.9/M$1/M▲ 11.1%
- Output price$1.80/M$2/M▲ 11.1%
- Cache read$0.09/M$0.1/M▲ 11.1%
via crossmodel
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via cortecs
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.12/M$0.11/M▼ 8.3%
- Output price$0.42/M$0.408/M▼ 2.9%
- cache write$0.06/M$0.055/M▼ 8.3%
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$24.75/M$27.50/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
- cache write$6.19/M$6.88/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via scnet-token-plan
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.148/M$0.165/M▲ 11.1%
- Output price$0.594/M$0.66/M▲ 11.1%
- Cache read$0.148/M$0.165/M▲ 11.1%
via vercel
What changed
First observed in the model catalog
via vancine
What changed
vancine began listing this model
via requesty
What changed
- Input price$0.126/M$0.14/M▲ 11.1%
- Output price$0.522/M$0.58/M▲ 11.1%
- Cache read$0.032/M$0.035/M▲ 11.1%
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.855/M$0.95/M▲ 11.1%
- Output price$3.60/M$4/M▲ 11.1%
- Cache read$0.855/M$0.95/M▲ 11.1%
via openrouter
What changed
- Output limit236K82K−65%
- Input price$0.6/M$0.32/M▼ 46.7%
- Output price$3.60/M$3.20/M▼ 11.1%
- Cache read$0.12/M—
via nano-gpt
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via orcarouter
What changed
- Context200K1M5×
via nano-gpt
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$10.80/M$12/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
- tiers[object Object][object Object]
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via neuralwatt
What changed
Removed from the model catalog
via orcarouter
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$22.50/M$25/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$5.63/M$6.25/M▲ 11.1%
via requesty
What changed
- Input price$0.148/M$0.165/M▲ 11.1%
- Output price$0.594/M$0.66/M▲ 11.1%
- Cache read$0.148/M$0.165/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via kilo
What changed
- Context1.05M1.05M1×
- Output limit944K384K−59%
via requesty
What changed
- Input price$0.36/M$0.4/M▲ 11.1%
- Output price$1.80/M$2/M▲ 11.1%
- Cache read$0.09/M$0.1/M▲ 11.1%
via kilo
What changed
- Output price$0.696/M$0.8/M▲ 14.9%
via openreason
What changed
openreason began listing this model
via openrouter
What changed
- Output limit262K131K−50%
- Cache read$0.25/M$0.2/M▼ 20.0%
via orcarouter
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via kilo
What changed
- Context32K33K1×
- Output limit26K26K1×
via tokengo
What changed
tokengo began listing this model
via neuralwatt
What changed
Removed from the model catalog
via requesty
What changed
- Input price$1.57/M$1.75/M▲ 11.1%
- Output price$12.60/M$14/M▲ 11.1%
- Cache read$0.158/M$0.175/M▲ 11.1%
via openrouter
What changed
- Input price$0.083/M$0.132/M▲ 60.0%
- Output price$0.33/M$0.528/M▲ 60.0%
- Cache read$0.021/M$0.033/M▲ 60.0%
via tokengo
What changed
tokengo began listing this model
via openrouter
What changed
- Input price$0.425/M$0.4/M▼ 5.9%
- Cache read$0.085/M$0.05/M▼ 41.2%
- cache write$0.531/M—
via ofox
What changed
- structured output—YesEnabled
via requesty
What changed
- Input price$0.054/M$0.06/M▲ 11.1%
- Output price$0.216/M$0.24/M▲ 11.1%
- Cache read$0.054/M$0.06/M▲ 11.1%
via volcengine-coding-plan
What changed
volcengine-coding-plan began listing this model
via openrouter
What changed
- Output limit118K16K−86%
- Input price$0.35/M$0.3/M▼ 14.3%
- Output price$1.50/M$1.20/M▼ 20.0%
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$5.40/M$6/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
via nano-gpt
What changed
First observed in the model catalog
via fireworks-ai
What changed
Removed from the model catalog
via crof
What changed
- Input price$0.25/M$0.2/M▼ 20.0%
- Output price$2.10/M$1.50/M▼ 28.6%
- Cache read$0.06/M$0.03/M▼ 50.0%
via requesty
What changed
- Input price$1.26/M$1.40/M▲ 11.1%
- Output price$3.96/M$4.40/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via fireworks-ai
What changed
Removed from the model catalog
via digitalocean
What changed
- Input price$0.87/M$1.74/M▲ 100.0%
- Output price$1.74/M$3.48/M▲ 100.0%
- Cache read$0.174/M$0.348/M▲ 100.0%
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$27/M$30/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- tiers[object Object][object Object]
via requesty
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.408/M$0.404/M▼ 1.0%
- Output price$1.51/M$1.50/M▼ 1.1%
- cache write$0.204/M$0.202/M▼ 1.0%
via digitalocean
What changed
- Input price$2.85/M$3/M▲ 5.3%
- Output price$14.25/M$15/M▲ 5.3%
- Cache read$0.285/M$0.3/M▲ 5.3%
via kilo
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via openrouter
What changed
- family—Hy
- open weightsNoYesEnabled
via tokengo
What changed
tokengo began listing this model
via edenai
What changed
First observed in the model catalog
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$27/M$30/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- tiers[object Object][object Object]
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$2.97/M$3.30/M▲ 11.1%
- Output price$14.85/M$16.50/M▲ 11.1%
- Cache read$0.27/M$0.3/M▲ 11.1%
- cache write$3.71/M$4.13/M▲ 11.1%
via neuralwatt
What changed
Removed from the model catalog
via requesty
What changed
- Input price$1.26/M$1.40/M▲ 11.1%
- Output price$3.96/M$4.40/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via github-copilot
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
via alibaba-token-plan
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$5.40/M$6/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$1.80/M$2/M▲ 11.1%
- +1 more changes
via huggingface
What changed
First observed in the model catalog
via digitalocean
What changed
- Input price$4/M$5/M▲ 25.0%
- Output price$20/M$30/M▲ 50.0%
- Cache read$0.4/M$0.5/M▲ 25.0%
- tiers[object Object][object Object]
via kilo
What changed
- Output limit131K236K1.8×
via volcengine
What changed
- structured output—YesEnabled
via orcarouter
What changed
Removed from the model catalog
via volcengine
What changed
- structured output—YesEnabled
via alibaba-cn
What changed
First observed in the model catalog
via vercel
What changed
- Context1.05M1M−5%
- Output limit1.05M384K−63%
- Input price$1.74/M$0.66/M▼ 62.1%
- Output price$3.48/M$1.98/M▼ 43.1%
- +1 more changes
via requesty
What changed
- Input price$0.126/M$0.14/M▲ 11.1%
- Output price$0.252/M$0.28/M▲ 11.1%
- Cache read$0.063/M$0.07/M▲ 11.1%
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$5.40/M$6/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
- cache write$1.80/M$2/M▲ 11.1%
- +1 more changes
via digitalocean
What changed
- Input price$0.25/M$0.5/M▲ 100.0%
- Output price$0.8/M$1.60/M▲ 100.0%
- Cache read$0.075/M$0.15/M▲ 100.0%
via orcarouter
What changed
- Cache read$0.06/M$0.03/M▼ 50.0%
via kilo
What changed
- Context131K262K2×
- Output limit33K16K−50%
via neuralwatt
What changed
Removed from the model catalog
via nvidia
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.36/M$0.4/M▲ 11.1%
- Output price$2.70/M$3/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via requesty
What changed
- Input price$1.57/M$1.75/M▲ 11.1%
- Output price$3.15/M$3.50/M▲ 11.1%
- Cache read$0.396/M$0.44/M▲ 11.1%
via requesty
What changed
- Input price$0.396/M$0.44/M▲ 11.1%
- Output price$1.98/M$2.20/M▲ 11.1%
- Cache read$0.396/M$0.44/M▲ 11.1%
via requesty
What changed
- Input price$1.13/M$1.25/M▲ 11.1%
- Output price$4.05/M$4.50/M▲ 11.1%
- Cache read$0.279/M$0.31/M▲ 11.1%
via kilo
What changed
- nameTencent: Hy4 previewHy4 preview
- family—Hy
- open weightsNoYesEnabled
via tokengo
What changed
tokengo began listing this model
via digitalocean
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$2.25/M$2.50/M▲ 11.1%
- Cache read$0.068/M$0.075/M▲ 11.1%
- cache write$0.495/M$0.55/M▲ 11.1%
via requesty
What changed
- Input price$3.60/M$4/M▲ 11.1%
- Output price$18/M$20/M▲ 11.1%
- Cache read$0.36/M$0.4/M▲ 11.1%
- cache write$4.50/M$5/M▲ 11.1%
- +1 more changes
via requesty
What changed
- Input price$0.27/M$0.3/M▲ 11.1%
- Output price$2.25/M$2.50/M▲ 11.1%
via cortecs
What changed
First observed in the model catalog
via tencent-tokenhub
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via requesty
What changed
- Input price$1.13/M$1.25/M▲ 11.1%
- Output price$2.25/M$2.50/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
- cache write$1.13/M$1.25/M▲ 11.1%
- +1 more changes
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via neuralwatt
What changed
- Cache read$0.036/M$0.014/M▼ 60.0%
via nano-gpt
What changed
First observed in the model catalog
via vancine
What changed
vancine began listing this model
via kilo
What changed
- Output limit147K66K−56%
via orcarouter
What changed
- Input price$2.50/M$1.25/M▼ 50.0%
- Output price$15/M$10/M▼ 33.3%
via requesty
What changed
- Input price$0.063/M$0.07/M▲ 11.1%
- Output price$0.306/M$0.34/M▲ 11.1%
- Cache read$0.063/M$0.07/M▲ 11.1%
via tokengo
What changed
tokengo began listing this model
via tokengo
What changed
tokengo began listing this model
via wandb
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via runinfra
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit33K16K−50%
- Input price$0.13/M$0.15/M▲ 15.4%
- Output price$0.52/M$0.6/M▲ 15.4%
via vercel
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via orcarouter
What changed
- Cache read—$0.0075/M
via inceptron
What changed
- Input price$0.75/M$0.71/M▼ 5.3%
- Output price$2.40/M$2.35/M▼ 2.1%
- Cache read$0.17/M$0.12/M▼ 29.4%
via wandb
What changed
First observed in the model catalog
via volcengine
What changed
- structured output—YesEnabled
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via ofox
What changed
- structured output—YesEnabled
via kenari
What changed
First observed in the model catalog
via fireworks-ai
What changed
First observed in the model catalog
via ofox
What changed
- structured output—YesEnabled
via digitalocean
What changed
- Input price$0.2/M$0.25/M▲ 25.0%
- Output price$0.696/M$0.87/M▲ 25.0%
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$10.80/M$12/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
- cache write$4.05/M$4.50/M▲ 11.1%
- +1 more changes
via requesty
What changed
- Input price$4.50/M$5/M▲ 11.1%
- Output price$22.50/M$25/M▲ 11.1%
- Cache read$0.45/M$0.5/M▲ 11.1%
- cache write$5.63/M$6.25/M▲ 11.1%
via orcarouter
What changed
Removed from the model catalog
via requesty
What changed
- Input price$0.28/M$0.32/M▲ 14.3%
- Output price$1.12/M$1.28/M▲ 14.3%
- Cache read$0.028/M$0.032/M▲ 14.3%
- cache write$0.35/M$0.4/M▲ 14.3%
via nano-gpt
What changed
- Input price$0.1/M$0.2/M▲ 100.0%
- Output price$0.6/M$1.20/M▲ 100.0%
- Cache read$0.01/M$0.02/M▲ 100.0%
- cache write$0.125/M$0.25/M▲ 100.0%
via orcarouter
What changed
Removed from the model catalog
via ollama-cloud
What changed
First observed in the model catalog
via baseten
What changed
First observed in the model catalog
via fireworks-ai
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.0%
- Output price$0.757/M$0.757/M▼ 0.0%
via fireworks-ai
What changed
Removed from the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- Input price$2.48/M$2.75/M▲ 11.1%
- Output price$14.85/M$16.50/M▲ 11.1%
- Cache read$0.247/M$0.275/M▲ 11.1%
via requesty
What changed
- Input price$0.54/M$0.6/M▲ 11.1%
- Output price$2.16/M$2.40/M▲ 11.1%
- Cache read$0.108/M$0.12/M▲ 11.1%
via ofox
What changed
- structured output—YesEnabled
via orcarouter
What changed
First observed in the model catalog
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.247/M$0.275/M▲ 11.1%
- Output price$1.49/M$1.65/M▲ 11.1%
- Cache read$0.025/M$0.028/M▲ 11.1%
- cache write$0.082/M$0.092/M▲ 11.1%
- +1 more changes
via requesty
What changed
- Input price$1.50/M$1.87/M▲ 25.0%
- Output price$3.74/M$4.68/M▲ 25.0%
- Cache read$0.299/M$0.374/M▲ 25.0%
via orcarouter
What changed
First observed in the model catalog
via regolo-ai
What changed
Removed from the model catalog
via github-copilot
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
- +1 more changes
via requesty
What changed
- Input price$27/M$30/M▲ 11.1%
- Output price$162/M$180/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via modal
What changed
First observed in the model catalog
via volcengine
What changed
- structured output—YesEnabled
via requesty
What changed
- Input price$0.45/M$0.5/M▲ 11.1%
- Output price$2.25/M$2.50/M▲ 11.1%
via kilo
What changed
- Output limit118K16K−86%
via tokengo
What changed
tokengo began listing this model
via requesty
What changed
- Input price$1.80/M$2/M▲ 11.1%
- Output price$9/M$10/M▲ 11.1%
- Cache read$0.18/M$0.2/M▲ 11.1%
- cache write$2.25/M$2.50/M▲ 11.1%
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$24.75/M$27.50/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
- cache write$6.19/M$6.88/M▲ 11.1%
via requesty
What changed
- Input price$0.126/M$0.14/M▲ 11.1%
- Output price$0.9/M$1/M▲ 11.1%
- Cache read$0.045/M$0.05/M▲ 11.1%
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$24.75/M$27.50/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
- cache write$6.19/M$6.88/M▲ 11.1%
via kilo
What changed
- Output limit236K82K−65%
via requesty
What changed
- Input price$0.018/M$0.02/M▲ 11.1%
- Output price$0.09/M$0.1/M▲ 11.1%
via opencode
What changed
First observed in the model catalog
via orcarouter
What changed
- Input price$1.50/M$0.5/M▼ 66.7%
- Output price$9/M$3/M▼ 66.7%
- Cache read$0.15/M$0.1/M▼ 33.3%
- input audio$1.50/M—
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.234/M$0.26/M▲ 11.1%
- Output price$2.34/M$2.60/M▲ 11.1%
via nano-gpt
What changed
First observed in the model catalog
via requesty
What changed
- Input price$2.70/M$3/M▲ 11.1%
- Output price$13.50/M$15/M▲ 11.1%
- Cache read$0.27/M$0.3/M▲ 11.1%
- cache write$3.38/M$3.75/M▲ 11.1%
- +1 more changes
via hyper
What changed
- Input price$0.274/M$0.284/M▲ 3.6%
- Output price$0.899/M$0.934/M▲ 3.9%
- cache write$0.137/M$0.142/M▲ 3.6%
via requesty
What changed
- Input price$0.045/M$0.05/M▲ 11.1%
- Output price$0.18/M$0.2/M▲ 11.1%
- Cache read$0.0090/M$0.01/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$2.70/M$3/M▲ 11.1%
- Output price$13.50/M$15/M▲ 11.1%
- Cache read$0.27/M$0.3/M▲ 11.1%
- cache write$3.38/M$3.75/M▲ 11.1%
via vercel
What changed
- Output price$0.15/M$0.2/M▲ 33.3%
- Cache read$0.05/M$0.01/M▼ 80.0%
via orcarouter
What changed
- Input price$0.19/M$0.147/M▼ 22.6%
- Output price$0.37/M$0.295/M▼ 20.3%
- Cache read$0.0028/M$0.02/M▲ 614.3%
via requesty
What changed
- Input price$1.57/M$1.75/M▲ 11.1%
- Output price$12.60/M$14/M▲ 11.1%
- Cache read$0.158/M$0.175/M▲ 11.1%
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$1.08/M$1.20/M▲ 11.1%
- Output price$3.78/M$4.20/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via requesty
What changed
- Input price$0.45/M$0.5/M▲ 11.1%
- Output price$2.70/M$3/M▲ 11.1%
- Cache read$0.09/M$0.1/M▲ 11.1%
via requesty
What changed
- Input price$0.054/M$0.06/M▲ 11.1%
- Output price$0.216/M$0.24/M▲ 11.1%
- Cache read$0.054/M$0.06/M▲ 11.1%
via kenari
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via ofox
What changed
- structured output—YesEnabled
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$29.70/M$33/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via requesty
What changed
- Input price$4.95/M$5.50/M▲ 11.1%
- Output price$24.75/M$27.50/M▲ 11.1%
- Cache read$0.495/M$0.55/M▲ 11.1%
- cache write$6.19/M$6.88/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via digitalocean
What changed
- Input price$0.08/M$0.14/M▲ 75.0%
- Output price$0.252/M$0.28/M▲ 11.1%
- Cache read$0.025/M$0.028/M▲ 11.1%
via orcarouter
What changed
First observed in the model catalog
via crossmodel
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.075/M$0.089/M▲ 18.3%
- Output price$0.15/M$0.177/M▲ 18.3%
- Cache read$0.015/M$0.018/M▲ 18.3%
via orcarouter
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit236K82K−65%
- Input price$0.26/M$0.29/M▲ 11.5%
- Output price$2.08/M$2.40/M▲ 15.4%
via ofox
What changed
- structured output—YesEnabled
via orcarouter
What changed
- Input price$60/M$30/M▼ 50.0%
- Output price$270/M$180/M▼ 33.3%
via requesty
What changed
- Input price$1.09/M$1.21/M▲ 11.1%
- Output price$4.36/M$4.84/M▲ 11.1%
- Cache read$0.272/M$0.302/M▲ 11.1%
via requesty
What changed
- Input price$2.25/M$2.50/M▲ 11.1%
- Output price$6.75/M$7.50/M▲ 11.1%
- Cache read$0.225/M$0.25/M▲ 11.1%
- cache write$2.81/M$3.13/M▲ 11.1%
via kenari
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.375/M$0.75/M▲ 100.0%
- Output price$1.88/M$3.75/M▲ 100.0%
- Cache read$0.037/M$0.075/M▲ 100.0%
- cache write$0.021/M$0.042/M▲ 100.0%
- +1 more changes
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via vercel
What changed
- Output limit131K128K−2%
- Cache read$0.2/M$0.25/M▲ 25.0%
via requesty
What changed
- Input price$0.135/M$0.15/M▲ 11.1%
- Output price$0.45/M$0.5/M▲ 11.1%
- Cache read$0.027/M$0.03/M▲ 11.1%
via hyper
What changed
- Input price$1.29/M$1.31/M▲ 1.9%
- Output price$4.22/M$4.27/M▲ 1.1%
- cache write$0.645/M$0.657/M▲ 1.9%
via requesty
What changed
- Input price$0.396/M$0.44/M▲ 11.1%
- Output price$1.19/M$1.32/M▲ 11.1%
- Cache read$0.013/M$0.014/M▲ 11.1%
via kenari
What changed
First observed in the model catalog
via requesty
What changed
- Input price$0.05/M$0.055/M▲ 11.1%
- Output price$0.396/M$0.44/M▲ 11.1%
- Cache read$0.0050/M$0.0055/M▲ 11.1%
via nano-gpt
What changed
- Input price$1.40/M$1/M▼ 28.6%
- Output price$4.40/M$3.20/M▼ 27.3%
- Cache read$0.26/M$0.2/M▼ 23.1%
via orcarouter
What changed
- Input price$4/M$2/M▼ 50.0%
- Output price$18/M$12/M▼ 33.3%
- input audio—$2/M
via neuralwatt
What changed
Removed from the model catalog
via volcengine
What changed
- structured output—YesEnabled
via llmgateway
What changed
- Input price$0.15/M$0.13/M▼ 13.3%
- Output price$0.5/M$0.4/M▼ 20.0%
- Cache read$0.03/M$0.024/M▼ 20.0%
via requesty
What changed
- Input price$0.396/M$0.44/M▲ 11.1%
- Output price$1.58/M$1.76/M▲ 11.1%
- Cache read$0.099/M$0.11/M▲ 11.1%
via requesty
What changed
- Input price$1.08/M$1.20/M▲ 11.1%
- Output price$3.78/M$4.20/M▲ 11.1%
- Cache read$0.234/M$0.26/M▲ 11.1%
via requesty
What changed
- Input price$0.099/M$0.11/M▲ 11.1%
- Output price$0.396/M$0.44/M▲ 11.1%
- Cache read$0.025/M$0.028/M▲ 11.1%
via edenai
What changed
- Input price$1.05/M$1.05/M▼ 0.0%
- Output price$1.05/M$1.05/M▼ 0.0%
via wandb
What changed
Removed from the model catalog
via edenai
What changed
Removed from the model catalog
via tokengo
What changed
tokengo began listing this model
via volcengine
What changed
- reasoningNoYesEnabled
via orcarouter
What changed
First observed in the model catalog
via kilo
What changed
- structured outputNoYesEnabled
- Output limit131K944K7.2×
- Input price$1.40/M$1.20/M▼ 14.3%
- Output price$4.40/M$1.20/M▼ 72.7%
- +1 more changes
via tokengo
What changed
tokengo began listing this model
via kilo
What changed
- Context1M262K−74%
via neuralwatt
What changed
Removed from the model catalog
via edenai
What changed
- Input price$0.757/M$0.757/M▼ 0.0%
- Output price$0.757/M$0.757/M▼ 0.0%
via openrouter
What changed
- Input price$0.87/M$0.772/M▼ 11.3%
- Output price$1.74/M$1.54/M▼ 11.3%
- Cache read$0.072/M$0.064/M▼ 11.3%
via requesty
What changed
- Input price$0.18/M$0.2/M▲ 11.1%
- Output price$1.13/M$1.25/M▲ 11.1%
- Cache read$0.018/M$0.02/M▲ 11.1%
via requesty
What changed
- Input price$1.49/M$1.65/M▲ 11.1%
- Output price$8.91/M$9.90/M▲ 11.1%
- Cache read$0.148/M$0.165/M▲ 11.1%
- cache write$1.57/M$1.74/M▲ 11.1%
via nano-gpt
What changed
Removed from the model catalog
via requesty
What changed
- Input price$1.98/M$2.20/M▲ 11.1%
- Output price$11.88/M$13.20/M▲ 11.1%
- Cache read$0.198/M$0.22/M▲ 11.1%
via crof
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
- tool callNoYesEnabled
- structured outputNoYesEnabled
via edenai
What changed
Removed from the model catalog
via orcarouter
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via ofox
What changed
- structured output—YesEnabled
via orcarouter
What changed
First observed in the model catalog
via orcarouter
What changed
First observed in the model catalog
via alibaba-cn
What changed
- Input price$0.86/M$0.573/M▼ 33.4%
- Output price$3.15/M$2.58/M▼ 18.1%
- tiers[object Object]
via vercel
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
First observed in the model catalog
via opper
What changed
- temperatureNoYesEnabled
via openai
What changed
- temperatureNoYesEnabled
via neuralwatt
What changed
First observed in the model catalog
via huggingface
What changed
First observed in the model catalog
via merge-gateway
What changed
- temperatureNoYesEnabled
via ofox
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Output limit16K236K14.4×
- Input price$0.09/M$0.087/M▼ 2.8%
- Output price$0.55/M$0.35/M▼ 36.4%
- Cache read—$0.018/M
via hyper
What changed
- Input price$0.188/M$0.178/M▼ 5.3%
- Output price$0.7/M$0.68/M▼ 2.9%
- cache write$0.094/M$0.089/M▼ 5.3%
via impossibl
What changed
- temperatureNoYesEnabled
via edenai
What changed
- Input price$0.758/M$0.757/M▼ 0.2%
- Output price$0.758/M$0.757/M▼ 0.2%
via requesty
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Input price$0.16/M$0.15/M▼ 6.3%
via volcengine
What changed
First observed in the model catalog
via databricks
What changed
- temperatureNoYesEnabled
via merge-gateway
What changed
- temperatureNoYesEnabled
via azure-cognitive-services
What changed
- temperatureNoYesEnabled
via vercel
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
First observed in the model catalog
via volcengine
What changed
First observed in the model catalog
via deepinfra
What changed
First observed in the model catalog
via openrouter
What changed
- structured outputNoYesEnabled
- Context1.05M1.31M1.3×
via orcarouter
What changed
- temperatureNoYesEnabled
via alibaba-cn
What changed
- Input price$0.87/M$0.825/M▼ 5.2%
- Output price$3.48/M$3.30/M▼ 5.1%
- tiers[object Object]
via nano-gpt
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Input price$0.14/M$0.1/M▼ 28.6%
- Output price$1/M$0.9/M▼ 10.0%
via baseten
What changed
First observed in the model catalog
via crossmodel
What changed
- temperatureNoYesEnabled
via edenai
What changed
- Input price$1.05/M$1.05/M▼ 0.2%
- Output price$1.05/M$1.05/M▼ 0.2%
via xpersona
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- temperatureNoYesEnabled
via github-copilot
What changed
- temperatureNoYesEnabled
via edenai
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Output price$0.075/M$0.1/M▲ 33.3%
via nearai
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- attachmentYesNoRemoved
- modalities.inputtext,image,video,pdftext
- Output limit66K131K2×
- Input price$0.25/M$1.40/M▲ 460.0%
- +2 more changes
via vercel
What changed
First observed in the model catalog
via fastrouter
What changed
- temperatureNoYesEnabled
via requesty
What changed
- temperatureNoYesEnabled
via hyper
What changed
- Context1.05M1M−5%
via merge-gateway
What changed
- temperatureNoYesEnabled
via ofox
What changed
- temperatureNoYesEnabled
via orcarouter
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
- temperatureNoYesEnabled
via kenari
What changed
- temperatureNoYesEnabled
via vercel
What changed
- structured output—YesEnabled
- release date2026-04-302026-04-17
- last updated2026-04-302026-04-17
via cloudflare-ai-gateway
What changed
- temperatureNoYesEnabled
via venice
What changed
- temperatureNoYesEnabled
via openrouter
What changed
Removed from the model catalog
via requesty
What changed
- temperatureNoYesEnabled
via venice
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- Cache read$0.033/M$0.035/M▲ 6.1%
via impossibl
What changed
- temperatureNoYesEnabled
via inceptron
What changed
- Input price$0.56/M$0.54/M▼ 3.6%
- Cache read$0.17/M$0.13/M▼ 23.5%
via vercel
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Input price$0.16/M$0.15/M▼ 6.3%
via kilo
What changed
- Output limit177K197K1.1×
- Cache read—$0/M
- reasoning—$0/M
via model-oracle-ai
What changed
- temperatureNoYesEnabled
via snowflake-cortex
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
First observed in the model catalog
via pioneer
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via ai-router
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Context512K262K−49%
- Output limit461K16K−96%
via azure
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via nano-gpt
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
via edenai
What changed
First observed in the model catalog
via requesty
What changed
- temperatureNoYesEnabled
via edenai
What changed
- temperatureNoYesEnabled
via kilo
What changed
First observed in the model catalog
via digitalocean
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- tiers[object Object][object Object]
via sap-ai-core
What changed
- temperatureNoYesEnabled
via vercel
What changed
First observed in the model catalog
via wandb
What changed
- Context128K131K1×
- Output limit128K131K1×
via fastrouter
What changed
- temperatureNoYesEnabled
via orcarouter
What changed
- temperatureNoYesEnabled
via orcarouter
What changed
- temperatureNoYesEnabled
via openai
What changed
- temperatureNoYesEnabled
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via tencent-token-plan
What changed
Removed from the model catalog
via hyper
What changed
First observed in the model catalog
via databricks
What changed
- temperatureNoYesEnabled
via tencent-tokenhub
What changed
Removed from the model catalog
via vercel
What changed
- temperatureYesNoRemoved
- release date2026-06-222026-05-30
- last updated2026-06-222026-05-30
via vercel
What changed
First observed in the model catalog
via unorouter
What changed
- temperatureNoYesEnabled
via neuralwatt
What changed
- Input price$0.104/M$0.14/M▲ 34.6%
- Output price$0.207/M$0.28/M▲ 35.3%
- Cache read$0.026/M$0.028/M▲ 7.7%
via venice
What changed
Removed from the model catalog
via azure-cognitive-services
What changed
- temperatureNoYesEnabled
via sap-ai-core
What changed
- temperatureNoYesEnabled
via ofox
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via openrouter
What changed
- Input price$0.132/M$0.083/M▼ 37.5%
- Output price$0.528/M$0.33/M▼ 37.5%
- Cache read$0.033/M$0.021/M▼ 37.5%
via abacus
What changed
- temperatureNoYesEnabled
via unorouter
What changed
- temperatureNoYesEnabled
via vercel
What changed
- temperatureNoYesEnabled
via venice
What changed
- temperatureNoYesEnabled
via databricks
What changed
- temperatureNoYesEnabled
via orcarouter
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via snowflake-cortex
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Input price$0.67/M$0.66/M▼ 1.5%
- Cache read$0.19/M$0.18/M▼ 5.3%
via hyper
What changed
- Input price$0.544/M$0.544/M▼ 0.1%
- Output price$2.85/M$2.76/M▼ 3.3%
- cache write$0.272/M$0.272/M▼ 0.1%
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
Removed from the model catalog
via openrouter
What changed
- Context512K262K−49%
- Output limit461K16K−96%
- Input price$0.6/M$0.5/M▼ 16.7%
- Output price$3.60/M$2.20/M▼ 38.9%
- +1 more changes
via alibaba
What changed
- tiers[object Object],[object Object]
via ofox
What changed
- temperatureNoYesEnabled
via edenai
What changed
- temperatureNoYesEnabled
via cloudflare-ai-gateway
What changed
- temperatureNoYesEnabled
via neuralwatt
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.2%
- Output price$0.7/M$0.699/M▼ 0.2%
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via orcarouter
What changed
- temperatureYesNoRemoved
via inceptron
What changed
- Input price$0.67/M$0.66/M▼ 1.5%
- Cache read$0.19/M$0.18/M▼ 5.3%
via edenai
What changed
- temperatureNoYesEnabled
via alibaba-cn
What changed
- tiers[object Object],[object Object]
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
- temperatureNoYesEnabled
via daoxe
What changed
- temperatureNoYesEnabled
via alibaba-cn
What changed
- Input price$0.43/M$0.172/M▼ 60.0%
- Output price$2.58/M$1.03/M▼ 60.0%
- reasoning$2.58/M$1.03/M▼ 60.0%
- tiers[object Object]
via unorouter
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Output limit82K236K2.9×
via github-copilot
What changed
- temperatureNoYesEnabled
via openrouter
What changed
First observed in the model catalog
via zhipuai
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.14/M$0.1/M▼ 28.6%
- Output price$1/M$0.9/M▼ 10.0%
via ofox
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
First observed in the model catalog
via opper
What changed
- temperatureNoYesEnabled
via merge-gateway
What changed
- Input price$0.075/M$0.015/M▼ 80.0%
- Output price$0.25/M$0.05/M▼ 80.0%
- Cache read$0.015/M$0.0030/M▼ 80.0%
via llmgateway
What changed
- Input price$0.076/M$0.051/M▼ 32.9%
- Output price$0.153/M$0.104/M▼ 32.0%
- Cache read$0.014/M$0.0097/M▼ 30.7%
via nano-gpt
What changed
- Context205K1.05M5.1×
- Input limit205K1.05M5.1×
- Input price$0.315/M$1.40/M▲ 344.4%
- Output price$1.26/M$4.40/M▲ 249.2%
- +1 more changes
via edenai
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Input price$1/M$0.435/M▼ 56.5%
- Output price$3/M$0.87/M▼ 71.0%
- Cache read$0.2/M$0.0040/M▼ 98.0%
via edenai
What changed
- Input price$0.467/M$0.466/M▼ 0.2%
- Output price$0.934/M$0.932/M▼ 0.2%
via empiriolabs
What changed
First observed in the model catalog
via alibaba
What changed
- tiers[object Object],[object Object]
via edenai
What changed
- temperatureNoYesEnabled
via github-copilot
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
- temperatureNoYesEnabled
via cloudflare-ai-gateway
What changed
- temperatureNoYesEnabled
via cortecs
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Input price$0.081/M$0.08/M▼ 1.9%
- Output price$0.162/M$0.159/M▼ 1.9%
- Cache read$0.016/M$0.016/M▼ 1.9%
via volcengine
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit236K228K−3%
- Cache read$0.03/M$0.025/M▼ 16.7%
via vercel
What changed
First observed in the model catalog
via nearai
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Output limit131K16K−88%
- Input price$0.5/M$0.43/M▼ 14.0%
- Output price$2/M$1.75/M▼ 12.5%
- Cache read$0.1/M$0.08/M▼ 20.0%
via opper
What changed
- temperatureNoYesEnabled
via abacus
What changed
- temperatureNoYesEnabled
via freemodel
What changed
- temperatureNoYesEnabled
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via alibaba-cn
What changed
- tiers[object Object],[object Object]
via hyper
What changed
- Input price$0.91/M$0.92/M▲ 1.1%
- Output price$2.81/M$2.98/M▲ 5.8%
- cache write$0.455/M$0.46/M▲ 1.1%
via openai
What changed
- temperatureNoYesEnabled
via kilo
What changed
- Context203K198K−2%
- Output limit131K16K−88%
- Input price$0.5/M$0.43/M▼ 14.0%
- Output price$2/M$1.75/M▼ 12.5%
- +1 more changes
via impossibl
What changed
- temperatureNoYesEnabled
via freemodel
What changed
- temperatureNoYesEnabled
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Output limit16K236K14.4×
via nano-gpt
What changed
Removed from the model catalog
via ofox
What changed
- temperatureNoYesEnabled
via orcarouter
What changed
- temperatureNoYesEnabled
via tencent-token-plan
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via kilo
What changed
- structured outputNoYesEnabled
via hyper
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$0.075/M$0.051/M▼ 32.0%
- Output price$0.175/M$0.104/M▼ 40.6%
- Cache read$0.015/M$0.0097/M▼ 37.4%
via nano-gpt
What changed
First observed in the model catalog
via azure-cognitive-services
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
First observed in the model catalog
via venice
What changed
- temperatureNoYesEnabled
via pioneer
What changed
- temperatureNoYesEnabled
via azure
What changed
- temperatureNoYesEnabled
via impossibl
What changed
- temperatureNoYesEnabled
via amazon-bedrock
What changed
- temperatureNoYesEnabled
via tencent-tokenhub
What changed
First observed in the model catalog
via freemodel
What changed
- temperatureNoYesEnabled
via vercel
What changed
- temperatureYesNoRemoved
via ofox
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via vercel
What changed
- structured output—YesEnabled
- release date2026-05-202026-04-16
- last updated2026-05-202026-04-16
via nearai
What changed
- temperatureNoYesEnabled
via anyapi
What changed
- temperatureNoYesEnabled
via abacus
What changed
- temperatureNoYesEnabled
via merge-gateway
What changed
- temperatureNoYesEnabled
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via cortecs
What changed
- temperatureNoYesEnabled
via hyper
What changed
- Input price$0.16/M$0.15/M▼ 6.3%
via openrouter
What changed
- Input price$0.06/M$0.05/M▼ 16.7%
- Output price$0.12/M$0.1/M▼ 16.7%
- Cache read$0.012/M$0.01/M▼ 16.7%
via nano-gpt
What changed
- attachmentYesNoRemoved
- modalities.inputtext,image,video,pdftext
- Output limit66K131K2×
- Input price$0.25/M$1.40/M▲ 460.0%
- +2 more changes
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via zai-coding-plan
What changed
First observed in the model catalog
via venice
What changed
- Input price$0.094/M$0.15/M▲ 60.0%
- Output price$0.313/M$0.5/M▲ 60.0%
- Cache read$0.019/M$0.03/M▲ 60.0%
via openrouter
What changed
- Output limit236K66K−72%
- Input price$0.25/M$0.225/M▼ 10.0%
- Output price$1.25/M$1.80/M▲ 44.0%
- Cache read$0.25/M$0.225/M▼ 10.0%
via zhipuai
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.2%
- Output price$0.758/M$0.757/M▼ 0.2%
via edenai
What changed
- temperatureNoYesEnabled
via volcengine
What changed
First observed in the model catalog
via snowflake-cortex
What changed
- temperatureNoYesEnabled
via fastrouter
What changed
- temperatureNoYesEnabled
via pioneer
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Output limit82K236K2.9×
- Input price$0.32/M$0.6/M▲ 87.5%
- Output price$3.20/M$3.60/M▲ 12.5%
- Cache read—$0.12/M
via opencode
What changed
- temperatureNoYesEnabled
via pioneer
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via openai
What changed
- temperatureNoYesEnabled
via ofox
What changed
- temperatureNoYesEnabled
via cloudflare-ai-gateway
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
- temperatureNoYesEnabled
via kilo
What changed
- modalities.inputtext,image,videotext,image
- Output limit944K1.05M1.1×
- Cache read—$0/M
- reasoning—$0/M
via nano-gpt
What changed
- attachmentYesNoRemoved
- modalities.inputtext,image,pdftext
- Context1M1.05M1×
- Output limit128K131K1×
- +4 more changes
via edenai
What changed
First observed in the model catalog
via xpersona
What changed
- temperatureNoYesEnabled
via abacus
What changed
- temperatureNoYesEnabled
via neuralwatt
What changed
First observed in the model catalog
via zhipuai-coding-plan
What changed
First observed in the model catalog
via abacus
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- attachmentYesNoRemoved
- modalities.inputtext,image,pdftext
- Context1M1.05M1×
- Output limit128K131K1×
- +4 more changes
via anyapi
What changed
- temperatureNoYesEnabled
via hyper
What changed
- Input price$0.598/M$0.638/M▲ 6.7%
- Output price$0.738/M$0.768/M▲ 4.1%
- cache write$0.299/M$0.319/M▲ 6.7%
via azure
What changed
- temperatureNoYesEnabled
via empiriolabs
What changed
First observed in the model catalog
via kilo
What changed
- Output limit236K228K−3%
via anyapi
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- temperatureNoYesEnabled
via databricks
What changed
- temperatureNoYesEnabled
via impossibl
What changed
- temperatureNoYesEnabled
via wandb
What changed
- Context128K131K1×
- Output limit128K131K1×
via vercel
What changed
- structured output—YesEnabled
- knowledge—2026-02-01
via zai
What changed
First observed in the model catalog
via nano-gpt
What changed
- Context205K1.05M5.1×
- Input limit205K1.05M5.1×
- Input price$0.315/M$1.40/M▲ 344.4%
- Output price$1.26/M$4.40/M▲ 249.2%
- +1 more changes
via databricks
What changed
- temperatureNoYesEnabled
via opper
What changed
- temperatureNoYesEnabled
via llmgateway
What changed
- temperatureNoYesEnabled
via openai
What changed
- temperatureNoYesEnabled
via hyper
What changed
- Input price$1.36/M$1.29/M▼ 5.1%
- Output price$4.40/M$4.22/M▼ 4.1%
- cache write$0.68/M$0.645/M▼ 5.1%
via openrouter
What changed
- Output price$0.075/M$0.1/M▲ 33.3%
via volcengine
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via merge-gateway
What changed
- temperatureNoYesEnabled
via kilo
What changed
First observed in the model catalog
via hyper
What changed
First observed in the model catalog
via crossmodel
What changed
- temperatureNoYesEnabled
via kilo
What changed
Removed from the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via requesty
What changed
- temperatureNoYesEnabled
via merge-gateway
What changed
- temperatureYesNoRemoved
via hyper
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context1.05M1M−5%
via kilo
What changed
- Output limit236K66K−72%
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.6/M$0.45/M▼ 25.0%
- Output price$3/M$2.25/M▼ 25.0%
- Cache read$0.1/M$0.07/M▼ 30.0%
via github-copilot
What changed
- temperatureNoYesEnabled
via wandb
What changed
- Context262K1.05M4×
- Output limit262K1.05M4×
via nearai
What changed
- temperatureNoYesEnabled
via impossibl
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- temperatureNoYesEnabled
via github-copilot
What changed
- temperatureNoYesEnabled
via requesty
What changed
- temperatureNoYesEnabled
via kilo
What changed
Removed from the model catalog
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via venice
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- temperatureNoYesEnabled
via vercel
What changed
- structured output—YesEnabled
via abacus
What changed
- temperatureNoYesEnabled
via nano-gpt
What changed
- temperatureNoYesEnabled
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via model-oracle-ai
What changed
- temperatureNoYesEnabled
via cloudflare-workers-ai
What changed
First observed in the model catalog
via nearai
What changed
- temperatureNoYesEnabled
via openai
What changed
- temperatureNoYesEnabled
via hyper
What changed
- Input price$0.41/M$0.424/M▲ 3.4%
- Output price$1.52/M$1.61/M▲ 6.1%
- cache write$0.205/M$0.212/M▲ 3.4%
via model-oracle-ai
What changed
- temperatureNoYesEnabled
via crossmodel
What changed
- temperatureNoYesEnabled
via pioneer
What changed
- temperatureNoYesEnabled
via llmgateway-providers
What changed
- temperatureNoYesEnabled
via edenai
What changed
First observed in the model catalog
via merge-gateway
What changed
- Output limit128K512K4×
via amazon-bedrock
What changed
- Input price$5.50/M$4.40/M▼ 20.0%
- Output price$33/M$22/M▼ 33.3%
- Cache read$0.55/M$0.44/M▼ 20.0%
- cache write$6.88/M$5.50/M▼ 20.0%
- +1 more changes
via opencode-go
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit236K66K−72%
via wandb
What changed
Removed from the model catalog
via amazon-bedrock
What changed
- open weightsNoYesEnabled
via volcengine
What changed
volcengine began listing this model
via kilo
What changed
- Context128K32K−75%
- Output limit115K29K−75%
- Input price$0.8/M$0.25/M▼ 68.8%
- Output price$1/M$0.75/M▼ 25.0%
- +1 more changes
via edenai
What changed
- Input price$0.176/M$0.352/M▲ 100.0%
- Output price$0.528/M$1.06/M▲ 100.0%
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via opper
What changed
- Output limit128K512K4×
via kilo
What changed
- Input price$0.435/M$1/M▲ 129.9%
- Output price$0.87/M$3/M▲ 244.8%
- Cache read$0.0040/M$0.2/M▲ 4900.0%
via edenai
What changed
- Cache read—$0.15/M
- cache write—$0.15/M
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via merge-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.077/M$0.089/M▲ 15.1%
- Output price$0.154/M$0.177/M▲ 15.1%
- Cache read$0.015/M$0.018/M▲ 15.1%
via huggingface
What changed
First observed in the model catalog
via opencode-go
What changed
- status—deprecated
via gmicloud
What changed
First observed in the model catalog
via wandb
What changed
Removed from the model catalog
via llmgateway
What changed
- Output limit128K512K4×
via cortecs
What changed
- Input price$1.67/M$1.39/M▼ 16.6%
- Output price$5.57/M$7.13/M▲ 28.0%
- Cache read—$0.139/M
via llmgateway-providers
What changed
- Output limit128K512K4×
via nano-gpt
What changed
- Context512K1.05M2×
via volcengine
What changed
volcengine began listing this model
via scnet-token-plan
What changed
- Context512K1.05M2×
- Output limit128K512K4×
via edenai
What changed
- Input price$0.175/M$0.175/M▲ 0.1%
- Output price$0.758/M$0.758/M▲ 0.1%
via edenai
What changed
- Input price$0.758/M$0.758/M▲ 0.1%
- Output price$0.758/M$0.758/M▲ 0.1%
via jalapeno
What changed
- Output limit128K512K4×
via llmgateway-providers
What changed
- Context197K205K1×
via ofox
What changed
- Context200K66K−67%
- Output limit131K2K−98%
via opencode
What changed
- status—deprecated
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via amazon-bedrock
What changed
- open weightsNoYesEnabled
via ofox
What changed
- Context197K205K1×
- Output limit128K131K1×
via nano-gpt
What changed
First observed in the model catalog
via wandb
What changed
Removed from the model catalog
via gmicloud
What changed
First observed in the model catalog
via volcengine
What changed
volcengine began listing this model
via volcengine
What changed
volcengine began listing this model
via llmgateway-providers
What changed
- Context512K1.05M2×
via ofox
What changed
- Context512K1.05M2×
- Output limit128K512K4×
via zenmux
What changed
- Context512K1.05M2×
- Output limit128K512K4×
via openrouter
What changed
- Output limit16K115K7×
- Input price$0.1/M$0.71/M▲ 610.0%
- Output price$0.32/M$0.71/M▲ 121.9%
- Cache read—$0.71/M
via minimax-cn-coding-plan
What changed
- Context1M1.05M1×
- Output limit128K512K4×
via kilo
What changed
- Output limit236K16K−93%
- Input price$0.08/M$0.09/M▲ 12.5%
- Output price$0.35/M$0.34/M▼ 2.9%
- Cache read$0.01/M$0.05/M▲ 400.0%
via minimax-cn
What changed
- Context197K205K1×
- Output limit128K131K1×
via edenai
What changed
- Output limit128K131K1×
via nano-gpt
What changed
First observed in the model catalog
via huggingface
What changed
- Output limit128K131K1×
via nano-gpt
What changed
Removed from the model catalog
via amazon-bedrock
What changed
- structured outputYesNoRemoved
- Input price$0.22/M$0.2/M▼ 9.1%
- Output price$1.32/M$1.20/M▼ 9.1%
- Cache read$0.022/M$0.02/M▼ 9.1%
- +2 more changes
via edenai
What changed
- Input price$0.581/M$1.12/M▲ 93.2%
- Output price$1.74/M$3.37/M▲ 93.2%
via openrouter
What changed
- Output limit236K16K−93%
- Input price$0.1/M$0.09/M▼ 10.0%
- Cache read$0.1/M$0.05/M▼ 50.0%
via venice
What changed
- Input price$1.88/M$0.938/M▼ 50.0%
- Output price$9.38/M$4.69/M▼ 50.0%
- Cache read$0.188/M$0.094/M▼ 50.0%
via amazon-bedrock
What changed
- open weightsNoYesEnabled
via edenai
What changed
- Input price$1.05/M$1.05/M▲ 0.1%
- Output price$1.05/M$1.05/M▲ 0.1%
via deepinfra
What changed
- Output limit128K512K4×
via zai-coding-plan
What changed
First observed in the model catalog
via volcengine
What changed
volcengine began listing this model
via openrouter
What changed
- Input price$0.04/M$0.06/M▲ 50.0%
- Output price$0.08/M$0.12/M▲ 50.0%
- Cache read$0.0080/M$0.012/M▲ 50.0%
via minimax-cn-coding-plan
What changed
- Context197K205K1×
- Output limit128K131K1×
via wafer.ai
What changed
- Output limit128K512K4×
via huggingface
What changed
- Output limit128K512K4×
via minimax
What changed
- Context197K205K1×
- Output limit128K131K1×
via kilo
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via llmtech
What changed
llmtech began listing this model
via minimax-coding-plan
What changed
- Context1M1.05M1×
- Output limit128K512K4×
via kilo
What changed
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- Cache read$0.033/M$0.035/M▲ 6.1%
via minimax-coding-plan
What changed
- Context197K205K1×
- Output limit128K131K1×
via edenai
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via amazon-bedrock
What changed
- structured outputYesNoRemoved
- Input price$5.50/M$4/M▼ 27.3%
- Output price$33/M$20/M▼ 39.4%
- Cache read$0.55/M$0.4/M▼ 27.3%
- +2 more changes
via edenai
What changed
- Input price$0.466/M$0.467/M▲ 0.1%
- Output price$0.933/M$0.934/M▲ 0.1%
via openrouter
What changed
- Output limit115K29K−75%
- Input price$0.8/M$0.25/M▼ 68.8%
- Output price$1/M$0.75/M▼ 25.0%
- Cache read$0.4/M—
via hyper
What changed
- Context512K1.05M2×
via openrouter
What changed
- Input price$0.556/M$0.79/M▲ 42.2%
- Output price$1.11/M$1.58/M▲ 42.2%
- Cache read$0.046/M$0.066/M▲ 42.2%
via aixy
What changed
aixy began listing this model
via kilo
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit236K66K−72%
- Input price$0.5/M$0.39/M▼ 22.0%
- Output price$3.60/M$2.34/M▼ 35.0%
- Cache read$0.3/M—
via openrouter
What changed
- Output limit82K236K2.9×
- Input price$0.32/M$0.6/M▲ 87.5%
- Output price$3.20/M$3.60/M▲ 12.5%
- Cache read—$0.12/M
via kilo
What changed
First observed in the model catalog
via minimax-cn
What changed
- Context1M1.05M1×
- Output limit128K512K4×
via amazon-bedrock
What changed
- structured outputYesNoRemoved
- Input price$2.20/M$2/M▼ 9.1%
- Output price$13.20/M$12/M▼ 9.1%
- Cache read$0.22/M$0.2/M▼ 9.1%
- +2 more changes
via openrouter
What changed
- Input price$0.035/M$0.04/M▲ 14.3%
- Output price$0.28/M$0.08/M▼ 71.4%
via volcengine
What changed
volcengine began listing this model
via mistral
What changed
First observed in the model catalog
via venice
What changed
- Input price$3/M$2/M▼ 33.3%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.3/M$0.2/M▼ 33.3%
- cache write$3.75/M$2.50/M▼ 33.3%
via llmgateway
What changed
- Context197K205K1×
- Output limit128K131K1×
via kilo
What changed
- Context328K131K−60%
- Output limit16K8K−50%
via amazon-bedrock
What changed
- open weightsNoYesEnabled
via volcengine
What changed
volcengine began listing this model
via venice
What changed
- Input price$1.88/M$0.938/M▼ 50.0%
- Output price$9.38/M$4.69/M▼ 50.0%
- Cache read$0.188/M$0.094/M▼ 50.0%
via nano-gpt
What changed
- Context512K1.05M2×
via kilo
What changed
- nameOx AlphaOx Alpha (retires Aug 26)
via cline-pass
What changed
- Context512K1.05M2×
- Output limit128K512K4×
via kilo
What changed
- Input price$0.035/M$0.04/M▲ 14.3%
- Output price$0.36/M$0.08/M▼ 77.8%
via volcengine
What changed
volcengine began listing this model
via edenai
What changed
- Cache read—$0.07/M
- cache write—$0.07/M
via gmicloud
What changed
- Output limit128K512K4×
via edenai
What changed
- Context1M786K−21%
via kilo
What changed
- Context131K128K−2%
- Output limit16K115K7×
via wandb
What changed
Removed from the model catalog
via kilo
What changed
- Output limit82K236K2.9×
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via openrouter
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit16K8K−50%
- Input price$0.1/M$0.11/M▲ 10.0%
- Output price$0.3/M$0.34/M▲ 13.3%
- Cache read—$0.055/M
via openrouter
What changed
Removed from the model catalog
via volcengine
What changed
volcengine began listing this model
via inceptron
What changed
- Input price$0.57/M$0.56/M▼ 1.8%
via amazon-bedrock
What changed
- open weightsNoYesEnabled
via edenai
What changed
- structured outputNoYesEnabled
via zhipuai-coding-plan
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.175/M$0.175/M▲ 0.1%
- Output price$0.7/M$0.7/M▲ 0.1%
via merge-gateway
What changed
- Input price$4/M$5/M▲ 25.0%
- Output price$24/M$30/M▲ 25.0%
via amazon-bedrock
What changed
- open weightsNoYesEnabled
via requesty
What changed
- Output limit128K512K4×
via volcengine
What changed
volcengine began listing this model
via minimax
What changed
- Context1M1.05M1×
- Output limit128K512K4×
via edenai
What changed
- Output limit128K512K4×
via volcengine
What changed
volcengine began listing this model
via volcengine
What changed
volcengine began listing this model
via kenari
What changed
- Context512K1.05M2×
- Output limit128K512K4×
via openrouter
What changed
- Output limit512K461K−10%
via nano-gpt
What changed
First observed in the model catalog
via openai
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
- +1 more changes
via openrouter
What changed
- Output limit262K236K−10%
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit33K29K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Output limit262K236K−10%
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Context262K1.05M4×
via edenai
What changed
- tool callYesNoRemoved
via hyper
What changed
- Input price$1.33/M$1.36/M▲ 2.1%
- Output price$4.31/M$4.40/M▲ 2.0%
- cache write$0.666/M$0.68/M▲ 2.1%
via openrouter
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit60K54K−10%
via kilo
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit1.05M944K−10%
via kilo
What changed
- Output limit33K29K−10%
via kilo
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit131K118K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via kilo
What changed
- Output limit131K262K2×
via kilo
What changed
- Output limit1.05M944K−10%
via kilo
What changed
- Output limit500K450K−10%
via cloudflare-ai-gateway
What changed
- release date2026-04-142026-04-16
via openrouter
What changed
- Output limit500K450K−10%
via nano-gpt
What changed
- Cache read$0.01/M$0.05/M▲ 400.0%
via openrouter
What changed
- Output limit1M450K−55%
via kilo
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit13K52K4×
via edenai
What changed
- Input price$0.627/M$1.12/M▲ 78.9%
- Output price$1.88/M$3.37/M▲ 78.9%
via openrouter
What changed
- Output limit1.00M900K−10%
via edenai
What changed
- tool callYesNoRemoved
via kilo
What changed
- Output limit33K210K6.4×
via kilo
What changed
- Output limit4K4K−10%
via openai
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
- +1 more changes
via openrouter
What changed
- Output limit4K4K−10%
via kilo
What changed
- Output limit33K26K−20%
via openrouter
What changed
- Output limit131K105K−20%
via openrouter
What changed
- Output limit131K118K−10%
via evroc
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit131K105K−20%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
- Context1.05M1M−5%
- tiers[object Object]
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit256K230K−10%
via opencode
What changed
- Cache read$0.5/M$0.3/M▼ 40.0%
- tiers[object Object][object Object]
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via merge-gateway
What changed
- nameQwen3 Next 80B A3B InstructQwen3-Next 80B-A3B Instruct
via openrouter
What changed
- Output limit262K210K−20%
via pendra
What changed
pendra began listing this model
via cloudflare-ai-gateway
What changed
- Context400K128K−68%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit33K29K−10%
via nano-gpt
What changed
- Cache read$0.01/M$0.05/M▲ 400.0%
via openrouter
What changed
- Output limit131K262K2×
- Input price$0.966/M$1.19/M▲ 23.2%
- Output price$3.04/M$3.74/M▲ 23.2%
- Cache read$0.193/M$0.221/M▲ 14.4%
via openrouter
What changed
- Output limit128K102K−20%
via agnes
What changed
agnes began listing this model
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit33K118K3.6×
via openrouter
What changed
- Output limit66K59K−10%
via kilo
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit128K102K−20%
via pendra
What changed
pendra began listing this model
via kilo
What changed
- Output limit52K210K4×
via kilo
What changed
- Output limit52K210K4×
via kilo
What changed
- Output limit4K4K−10%
via kilo
What changed
- Output limit384K944K2.5×
via cloudflare-ai-gateway
What changed
- Context1.05M1M−5%
- Cache read$0.5/M—
- tiers[object Object]
via agnes
What changed
agnes began listing this model
via kilo
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit66K59K−10%
via cloudflare-ai-gateway
What changed
- Context400K128K−68%
via kilo
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit128K115K−10%
via nano-gpt
What changed
- Context131K262K2×
- Input limit131K262K2×
via openrouter
What changed
- Output limit128K115K−10%
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
- Context400K128K−68%
via kilo
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit1.05M944K−10%
via neosmith
What changed
neosmith began listing this model
via kilo
What changed
- Output limit6K26K4×
via kilo
What changed
- Output limit131K118K−10%
via nano-gpt
What changed
- Context66K262K4×
- Input limit66K262K4×
via openrouter
What changed
- Output limit262K236K−10%
via edenai
What changed
- Input price$0.4/M$0.44/M▲ 10.0%
- Output price$2/M$2.20/M▲ 10.0%
- Cache read—$0.044/M
via openrouter
What changed
- Output limit131K177K1.3×
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via llmgateway
What changed
- Context1M1.05M1×
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit33K29K−10%
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit128K203K1.6×
- Input price$0.966/M$1.26/M▲ 30.4%
- Output price$3.04/M$3.96/M▲ 30.4%
- Cache read$0.179/M$0.234/M▲ 30.4%
via kilo
What changed
- Context200K203K1×
- Output limit128K203K1.6×
via opencode-go
What changed
First observed in the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via nano-gpt
What changed
- Context66K262K4×
- Input limit66K262K4×
via kilo
What changed
- Output limit66K59K−10%
via openrouter
What changed
- Output limit66K52K−20%
via kilo
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit1.05M944K−10%
via openrouter
What changed
- Output limit256K230K−10%
via kilo
What changed
- Output limit131K118K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit1M900K−10%
via evroc
What changed
Removed from the model catalog
via kilo
What changed
- Output limit66K59K−10%
via kilo
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit33K26K−20%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Context1.05M164K−84%
via neosmith
What changed
neosmith began listing this model
via pendra
What changed
pendra began listing this model
via kilo
What changed
- Output price$0.12/M$0.075/M▼ 37.5%
- Cache read$0.01/M$0.0080/M▼ 20.0%
via openrouter
What changed
- Output limit1.05M944K−10%
via openrouter
What changed
- Output limit203K182K−10%
via kilo
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit197K177K−10%
via openrouter
What changed
- Output limit16K461K28.1×
via openrouter
What changed
- Output limit8K7K−10%
via standardcompute
What changed
standardcompute began listing this model
via openrouter
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit33K29K−10%
via openrouter
What changed
- Output limit1.05M944K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit262K210K−20%
via kilo
What changed
- Output limit262K236K−10%
via hyper
What changed
- Input price$0.86/M$0.91/M▲ 5.8%
- Output price$2.78/M$2.81/M▲ 1.1%
- cache write$0.43/M$0.455/M▲ 5.8%
via kilo
What changed
- Output limit26K118K4.5×
via kilo
What changed
- Output limit262K210K−20%
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via kilo
What changed
- Output limit256K230K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit33K29K−10%
via openrouter
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit131K118K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit33K29K−10%
via openrouter
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit16K7K−55%
via kilo
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
- release date2026-06-292026-06-30
via openrouter
What changed
- Input price$1/M$0.95/M▼ 5.0%
- Cache read$0.17/M$0.16/M▼ 5.9%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit4K4K−10%
via cline-pass
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit1.05M944K−10%
via neosmith
What changed
neosmith began listing this model
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via kilo
What changed
- Output limit512K461K−10%
via kilo
What changed
- Output limit1.00M900K−10%
via kilo
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit262K236K−10%
via edenai
What changed
- Input price$0.467/M$0.466/M▼ 0.0%
- Output price$0.933/M$0.933/M▼ 0.0%
via vercel
What changed
Removed from the model catalog
via pendra
What changed
pendra began listing this model
via openrouter
What changed
- Input price$0.056/M$0.089/M▲ 58.2%
- Output price$0.112/M$0.177/M▲ 58.2%
- Cache read$0.011/M$0.018/M▲ 58.2%
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit66K59K−10%
via openrouter
What changed
- Output limit262K210K−20%
via nano-gpt
What changed
- Context131K262K2×
- Input limit131K262K2×
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit164K147K−10%
via nano-gpt
What changed
- Context66K262K4×
- Input limit66K262K4×
via kilo
What changed
- Output limit26K102K4×
via kilo
What changed
- Output limit262K118K−55%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit164K147K−10%
via digitalocean
What changed
- Input price$1/M$2/M▲ 100.0%
- Output price$6/M$12/M▲ 100.0%
- Cache read$0.1/M$0.2/M▲ 100.0%
- tiers[object Object][object Object]
via kilo
What changed
- Output limit2M1.80M−10%
via kilo
What changed
- Output limit51K205K4×
via kilo
What changed
- Output limit26K115K4.5×
via kilo
What changed
- Output limit16K7K−55%
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit128K102K−20%
via openrouter
What changed
- Output limit131K105K−20%
via openrouter
What changed
- Output limit384K944K2.5×
via iteracompute
What changed
iteracompute began listing this model
via openrouter
What changed
- Output limit131K118K−10%
via hyper
What changed
- Input price$0.178/M$0.188/M▲ 5.6%
- Output price$0.68/M$0.7/M▲ 2.9%
- cache write$0.089/M$0.094/M▲ 5.6%
via openrouter
What changed
- Output limit262K210K−20%
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit1.05M944K−10%
via nano-gpt
What changed
- Context131K262K2×
- Input limit131K262K2×
via merge-gateway
What changed
- nameQwen3 Next 80B A3B ThinkingQwen3-Next 80B-A3B (Thinking)
via nano-gpt
What changed
- reasoningNoYesEnabled
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit25K114K4.5×
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Output limit131K944K7.2×
via deepseek
What changed
Removed from the model catalog
via openrouter
What changed
- Input price$0.4/M$0.425/M▲ 6.2%
- Output price$3/M$2.55/M▼ 15.0%
- Cache read$0.05/M$0.085/M▲ 70.0%
- cache write—$0.531/M
via cloudflare-ai-gateway
What changed
- tiers[object Object]
via openrouter
What changed
- Output limit128K102K−20%
via nano-gpt
What changed
- Context131K262K2×
- Input limit131K262K2×
via cloudflare-ai-gateway
What changed
- release date2026-06-072026-06-09
via inceptron
What changed
- Output price$2.90/M$2.40/M▼ 17.2%
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.0%
- Output price$0.758/M$0.758/M▼ 0.0%
via cloudflare-ai-gateway
What changed
- statusdeprecated—
- Context1.05M1M−5%
via cloudflare-ai-gateway
What changed
- statusdeprecated—
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via edenai
What changed
- tool callYesNoRemoved
via openrouter
What changed
- Output limit500K450K−10%
via kilo
What changed
- Context66K131K2×
via openrouter
What changed
- Output limit164K147K−10%
via cloudflare-ai-gateway
What changed
- Context400K128K−68%
via openrouter
What changed
- Output limit131K118K−10%
via pendra
What changed
pendra began listing this model
via cloudflare-ai-gateway
What changed
- statusdeprecated—
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit128K102K−20%
via cloudflare-ai-gateway
What changed
- Context1.05M1M−5%
- tiers[object Object]
via openrouter
What changed
- Output limit1.05M944K−10%
via kilo
What changed
- Context262K1M3.8×
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit1.05M944K−10%
via cloudflare-ai-gateway
What changed
- Context400K128K−68%
via kilo
What changed
Removed from the model catalog
via kilo
What changed
- Output limit33K105K3.2×
via openrouter
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit161K145K−10%
via openrouter
What changed
- Input price$0.066/M$0.14/M▲ 112.8%
- Output price$0.132/M$0.28/M▲ 112.8%
- Cache read$0.013/M$0.028/M▲ 112.8%
via kilo
What changed
- Output limit262K236K−10%
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit26K115K4.5×
via openrouter
What changed
- Output limit262K210K−20%
via digitalocean
What changed
- Input price$0.1/M$0.2/M▲ 100.0%
- Output price$0.6/M$1.20/M▲ 100.0%
- Cache read$0.01/M$0.02/M▲ 100.0%
- tiers[object Object][object Object]
via deepseek
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit66K59K−10%
via kilo
What changed
- Output limit1.05M944K−10%
via edenai
What changed
- Input price$0.176/M$0.352/M▲ 100.0%
- Output price$0.528/M$1.06/M▲ 100.0%
via cloudflare-ai-gateway
What changed
- tiers[object Object]
via openrouter
What changed
- Output limit66K59K−10%
via nano-gpt
What changed
- reasoningNoYesEnabled
via openrouter
What changed
Removed from the model catalog
via openrouter
What changed
- Output price$0.12/M$0.075/M▼ 37.5%
- Cache read$0.01/M$0.0080/M▼ 20.0%
via kilo
What changed
- Cache read—$0.1/M
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.522/M$0.79/M▲ 51.4%
- Output price$1.04/M$1.58/M▲ 51.4%
- Cache read$0.043/M$0.066/M▲ 51.4%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit262K236K−10%
via nano-gpt
What changed
First observed in the model catalog
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit1.05M944K−10%
via openrouter
What changed
- Output limit262K236K−10%
via kilo
What changed
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- Cache read$0.033/M$0.035/M▲ 6.1%
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Output limit33K26K−20%
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit262K236K−10%
via cline-pass
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit975K1.05M1.1×
- Input price$2.60/M$2.80/M▲ 7.7%
- Output price$13/M$14/M▲ 7.7%
via kilo
What changed
- Output limit262K236K−10%
via pendra
What changed
pendra began listing this model
via openrouter
What changed
- Output limit33K26K−20%
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via kilo
What changed
- Output limit8K7K−10%
via kilo
What changed
- Output limit512K461K−10%
via nano-gpt
What changed
Removed from the model catalog
via kilo
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit262K236K−10%
via digitalocean
What changed
- Input price$2/M$5/M▲ 150.0%
- Output price$10/M$30/M▲ 200.0%
- Cache read$0.2/M$0.5/M▲ 150.0%
- tiers[object Object][object Object]
via openrouter
What changed
- Output limit262K236K−10%
via opencode-go
What changed
- status—deprecated
- Cache read$0.5/M$0.3/M▼ 40.0%
- tiers[object Object][object Object]
via vivgrid
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit197K177K−10%
via openrouter
What changed
- Cache read—$0.1/M
via evroc
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit2M1.80M−10%
via openrouter
What changed
- Output limit262K210K−20%
via cloudflare-ai-gateway
What changed
- Input price$5/M$2/M▼ 60.0%
- Output price$30/M$10/M▼ 66.7%
- Cache read$0.5/M$0.25/M▼ 50.0%
- cache write$6.25/M$3.13/M▼ 50.0%
- +1 more changes
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit131K177K1.3×
via openrouter
What changed
- Output limit256K205K−20%
via kilo
What changed
- Output limit2M1.80M−10%
via openrouter
What changed
- Output limit262K236K−10%
via openrouter
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit262K210K−20%
via kilo
What changed
- Input price$0.03/M$0.02/M▼ 33.3%
- Output price$0.13/M$0.1/M▼ 23.1%
- Cache read$0.03/M—
via openrouter
What changed
- Output limit128K115K−10%
via openrouter
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via kilo
What changed
- Output limit164K147K−10%
via merge-gateway
What changed
- Input price$0.035/M$0.22/M▲ 528.6%
- Output price$0.07/M$0.66/M▲ 842.9%
via wandb
What changed
First observed in the model catalog
via kilo
What changed
- Cache read—$0.1/M
via llmgateway-providers
What changed
- Input price$1.69/M$1.32/M▼ 21.9%
- Output price$3.38/M$3.96/M▲ 17.2%
- Cache read$0.14/M$0.132/M▼ 5.7%
via kilo
What changed
- Output limit33K105K3.2×
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via kilo
What changed
- Output limit131K118K−10%
via hyper
What changed
- Input price$0.11/M$0.12/M▲ 9.1%
- Output price$0.408/M$0.42/M▲ 2.9%
- cache write$0.055/M$0.06/M▲ 9.1%
via kilo
What changed
- Output limit26K105K4×
via openrouter
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit262K236K−10%
via zai
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$0.14/M$0.44/M▲ 214.3%
- Output price$0.28/M$1.32/M▲ 371.4%
- Cache read$0.028/M$0.044/M▲ 57.1%
via openrouter
What changed
- Output limit2M1.80M−10%
via kilo
What changed
- Output limit4K4K−10%
via kilo
What changed
- Output limit66K59K−10%
via openrouter
What changed
- Context262K1.05M4×
via kilo
What changed
- Output limit4K900K219.7×
via kilo
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via kilo
What changed
- Output limit262K236K−10%
via cloudflare-ai-gateway
What changed
- Context1.05M1M−5%
- tiers[object Object]
via kilo
What changed
- Output limit1.05M944K−10%
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit262K236K−10%
via kilo
What changed
- Output limit500K450K−10%
via kilo
What changed
- Output limit131K118K−10%
via kilo
What changed
- Output limit500K450K−10%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via cloudflare-ai-gateway
What changed
- Context400K128K−68%
via kilo
What changed
- Output limit262K236K−10%
via vercel
What changed
- Context1M262K−74%
- Output limit33K131K4×
- Input price$0/M$0.05/M
- Output price$0/M$0.2/M
- +1 more changes
via kilo
What changed
- Output limit26K105K4×
via openrouter
What changed
- Output limit1.05M944K−10%
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via edenai
What changed
First observed in the model catalog
via kilo
What changed
- Output limit7K29K4.5×
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit32K26K−20%
via agnes
What changed
First observed in the model catalog
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via vercel
What changed
- Context131K128K−2%
- Output limit66K16K−76%
- Input limit66K112K1.7×
- Input price$0.075/M$0.07/M▼ 6.7%
- +2 more changes
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.0%
- Output price$0.7/M$0.7/M▼ 0.0%
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Output limit128K115K−10%
via openrouter
What changed
- Output limit131K118K−10%
via openrouter
What changed
- Output limit161K145K−10%
via openrouter
What changed
- Output limit262K210K−20%
via openrouter
What changed
- Output limit127K114K−10%
via edenai
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit33K29K−10%
via kilo
What changed
- Output limit256K230K−10%
via nano-gpt
What changed
- Context66K262K4×
- Input limit66K262K4×
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via edenai
What changed
- Input price$1.05/M$1.05/M▼ 0.0%
- Output price$1.05/M$1.05/M▼ 0.0%
via edenai
What changed
- Input price$0.758/M$0.758/M▼ 0.0%
- Output price$0.758/M$0.758/M▼ 0.0%
via cloudflare-ai-gateway
What changed
- release date2026-02-042026-02-05
via openrouter
What changed
- Output limit256K230K−10%
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Output limit60K54K−10%
via openrouter
What changed
- Output limit33K29K−10%
via openrouter
What changed
- Output limit16K15K−10%
via openrouter
What changed
- Output limit262K105K−60%
via openrouter
What changed
- Output limit4K4K−10%
via neosmith
What changed
neosmith began listing this model
via kilo
What changed
- Context975K1.05M1.1×
- Output limit975K1.05M1.1×
- Input price$2.60/M$2.80/M▲ 7.7%
- Output price$13/M$14/M▲ 7.7%
via cloudflare-ai-gateway
What changed
Removed from the model catalog
via openrouter
What changed
- Cache read—$0.1/M
via kilo
What changed
- Output limit16K15K−10%
via openrouter
What changed
Removed from the model catalog
via inceptron
What changed
- Output price$3.41/M$3.39/M▼ 0.6%
via cloudflare-ai-gateway
What changed
First observed in the model catalog
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via openrouter
What changed
- Output limit6K4K−33%
via alibaba-cn
What changed
- Cache read—$0.012/M
- cache write—$0.144/M
via openrouter
What changed
- Input price$0.24/M$0.3/M▲ 25.0%
- Output price$0.96/M$1.20/M▲ 25.0%
- Cache read$0.048/M$0.06/M▲ 25.0%
via edenai
What changed
- Input price$0.468/M$0.467/M▼ 0.3%
- Output price$0.936/M$0.933/M▼ 0.3%
via hyper
What changed
- Input price$1.36/M$1.33/M▼ 2.1%
- Output price$4.40/M$4.31/M▼ 2.0%
- cache write$0.68/M$0.666/M▼ 2.1%
via openrouter
What changed
- Output limit1.02M33K−97%
via openrouter
What changed
- Input price$0.45/M$0.6/M▲ 33.3%
- Output price$2.25/M$3/M▲ 33.3%
- Cache read$0.07/M$0.1/M▲ 42.9%
via kilo
What changed
- Input price$0.075/M$0.11/M▲ 46.7%
- Output price$0.2/M$0.33/M▲ 65.0%
- Cache read—$0.011/M
via hyper
What changed
- Context131K128K−2%
via edenai
What changed
- Input price$0.176/M$0.352/M▲ 100.0%
- Output price$0.528/M$1.06/M▲ 100.0%
via nano-gpt
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via inceptron
What changed
- Input price$0.6/M$0.57/M▼ 5.0%
via edenai
What changed
- Input price$1.05/M$1.05/M▼ 0.3%
- Output price$1.05/M$1.05/M▼ 0.3%
via runinfra
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
via edenai
What changed
- Input price$0.627/M$1.16/M▲ 85.3%
- Output price$1.88/M$3.48/M▲ 85.3%
via openrouter
What changed
- Output limit66K262K4×
- Input price$0.39/M$0.5/M▲ 28.2%
- Output price$2.34/M$3.60/M▲ 53.8%
- Cache read—$0.3/M
via edenai
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit8K16K2×
- Input price$0.13/M$0.12/M▼ 7.7%
- Output price$0.52/M$0.5/M▼ 3.8%
via kilo
What changed
- Context131K41K−69%
- Output limit8K16K2×
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via kilo
What changed
- Input price$0.5/M$0.425/M▼ 15.0%
- Output price$3/M$2.55/M▼ 15.0%
- Cache read$0.1/M$0.085/M▼ 15.0%
- cache write$0.625/M$0.531/M▼ 15.0%
via edenai
What changed
- Input price$0.76/M$0.758/M▼ 0.3%
- Output price$0.76/M$0.758/M▼ 0.3%
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.3%
- Output price$0.702/M$0.7/M▼ 0.3%
via vercel
What changed
Removed from the model catalog
via vercel
What changed
Removed from the model catalog
via hyper
What changed
- Context256K262K1×
via openrouter
What changed
- Input price$0.04/M$0.035/M▼ 12.5%
- Output price$0.08/M$0.13/M▲ 62.5%
- Cache read$0.0080/M$0.01/M▲ 25.0%
via openrouter
What changed
- Output limit262K82K−69%
- Input price$0.6/M$0.32/M▼ 46.7%
- Output price$3.60/M$3.20/M▼ 11.1%
- Cache read$0.12/M—
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Output limit66K262K4×
via digitalocean
What changed
- Input price$5/M$2/M▼ 60.0%
- Output price$30/M$10/M▼ 66.7%
- Cache read$0.5/M$0.2/M▼ 60.0%
- tiers[object Object][object Object]
via vercel
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit384K131K−66%
- Input price$0.08/M$0.14/M▲ 75.0%
- Output price$0.18/M$0.28/M▲ 55.6%
- Cache read$0.016/M$0.028/M▲ 75.0%
via kilo
What changed
- Input price$0.019/M$0.165/M▲ 768.4%
- Output price$0.03/M$0.165/M▲ 450.0%
- Cache read—$0.017/M
via kilo
What changed
- Context262K1.05M4×
via llmgateway
What changed
- Context262K1M3.8×
via llmgateway-providers
What changed
- Input price$6/M$1/M▼ 83.3%
- Output price$60/M$5/M▼ 91.7%
- Cache read$1.20/M$0.2/M▼ 83.3%
- cache write$7.50/M$1.25/M▼ 83.3%
via openrouter
What changed
- Input price$0.541/M$0.95/M▲ 75.4%
- Output price$2.28/M$4/M▲ 75.4%
- Cache read$0.091/M$0.16/M▲ 75.4%
via nano-gpt
What changed
Removed from the model catalog
via agentrouter
What changed
agentrouter began listing this model
via nano-gpt
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.479/M$0.44/M▼ 8.2%
- Output price$1.44/M$1.32/M▼ 8.2%
- Cache read$0.015/M$0.044/M▲ 188.7%
via openrouter
What changed
First observed in the model catalog
via vercel
What changed
- Input price$1.32/M$0.66/M▼ 50.0%
- Output price$3.96/M$1.98/M▼ 50.0%
- Cache read$0.132/M$0.066/M▼ 50.0%
via edenai
What changed
Removed from the model catalog
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.049/M$0.059/M▲ 20.1%
- Output price$0.098/M$0.117/M▲ 20.1%
- Cache read$0.0098/M$0.012/M▲ 20.1%
via llmgateway-providers
What changed
First observed in the model catalog
via llmgateway
What changed
- Input price$6/M$1/M▼ 83.3%
- Output price$60/M$5/M▼ 91.7%
- Cache read$1.20/M$0.2/M▼ 83.3%
- cache write$7.50/M$1.25/M▼ 83.3%
via kilo
What changed
- Output limit6K4K−33%
via openrouter
What changed
- Context262K131K−50%
via agentrouter
What changed
agentrouter began listing this model
via nano-gpt
What changed
Removed from the model catalog
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via vercel
What changed
First observed in the model catalog
via kilo
What changed
- Context262K1.05M4×
via edenai
What changed
- Input price$0.175/M$0.175/M▼ 0.3%
- Output price$0.76/M$0.758/M▼ 0.3%
via hyper
What changed
- Input price$0.638/M$0.598/M▼ 6.3%
- Output price$0.768/M$0.738/M▼ 3.9%
- cache write$0.319/M$0.299/M▼ 6.3%
via openrouter
What changed
Removed from the model catalog
via openrouter
What changed
Removed from the model catalog
via opencode-go
What changed
First observed in the model catalog
via llmgateway-providers
What changed
- Input price$3/M$1.20/M▼ 60.0%
- Output price$15/M$6/M▼ 60.0%
- Cache read$0.6/M$0.24/M▼ 60.0%
- cache write$3.75/M$1.50/M▼ 60.0%
via hyper
What changed
- Input price$0.188/M$0.178/M▼ 5.3%
- Output price$0.7/M$0.68/M▼ 2.9%
- cache write$0.094/M$0.089/M▼ 5.3%
via openrouter
What changed
- Cache read$0.17/M$0.19/M▲ 11.8%
via llmgateway-providers
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via nano-gpt
What changed
Removed from the model catalog
via inceptron
What changed
- Cache read$0.17/M$0.19/M▲ 11.8%
via edenai
What changed
- Context1M786K−21%
via kilo
What changed
- Output limit66K262K4×
via kilo
What changed
- Output limit384K131K−66%
via llmgateway-providers
What changed
- Context262K1M3.8×
- cache write—$0.625/M
via agentrouter
What changed
agentrouter began listing this model
via kilo
What changed
- Output limit262K82K−69%
via vercel
What changed
Removed from the model catalog
via chutes
What changed
- Input price$0.14/M$0.44/M▲ 214.3%
- Output price$0.28/M$1.32/M▲ 371.4%
- Cache read$0.014/M$0.044/M▲ 214.3%
via kilo
What changed
Removed from the model catalog
via digitalocean
What changed
- Input price$0.2/M$0.1/M▼ 50.0%
- Output price$1.20/M$0.6/M▼ 50.0%
- Cache read$0.02/M$0.01/M▼ 50.0%
- tiers[object Object][object Object]
via hyper
What changed
- Input price$0.408/M$0.41/M▲ 0.5%
- Output price$1.51/M$1.52/M▲ 0.5%
- cache write$0.204/M$0.205/M▲ 0.5%
via llmgateway-providers
What changed
- Input price$0.248/M$0.375/M▲ 51.2%
- Output price$1.49/M$2.25/M▲ 51.5%
via kilo
What changed
Removed from the model catalog
via kilo
What changed
- Input price$0.132/M$0.14/M▲ 6.1%
- Output price$0.528/M$0.58/M▲ 9.8%
- Cache read$0.033/M$0.035/M▲ 6.1%
via openrouter
What changed
Removed from the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via openrouter
What changed
- Input price$0.397/M$0.519/M▲ 30.8%
- Output price$0.794/M$1.04/M▲ 30.8%
- Cache read$0.033/M$0.043/M▲ 30.8%
via kilo
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit66K262K4×
via hyper
What changed
- Input price$0.91/M$0.86/M▼ 5.5%
- Output price$2.93/M$2.78/M▼ 5.1%
- cache write$0.455/M$0.43/M▼ 5.5%
via openrouter
What changed
- Input price$0.95/M$1/M▲ 5.3%
- Cache read$0.16/M$0.17/M▲ 6.3%
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.04/M$0.035/M▼ 12.5%
- Output price$0.08/M$0.13/M▲ 62.5%
- Cache read$0.0080/M$0.01/M▲ 25.0%
via nano-gpt
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via kilo
What changed
- Context1.02M33K−97%
- Output limit1.02M33K−97%
via edenai
What changed
Removed from the model catalog
via digitalocean
What changed
- Input price$2/M$1/M▼ 50.0%
- Output price$12/M$6/M▼ 50.0%
- Cache read$0.2/M$0.1/M▼ 50.0%
- tiers[object Object][object Object]
via openrouter
What changed
Removed from the model catalog
via kilo
What changed
Removed from the model catalog
via vercel
What changed
Removed from the model catalog
via opper
What changed
opper began listing this model
via edenai
What changed
- Input price$0.627/M$1.16/M▲ 85.3%
- Output price$1.88/M$3.48/M▲ 85.3%
via opper
What changed
opper began listing this model
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.541/M$0.95/M▲ 75.4%
- Output price$2.28/M$4/M▲ 75.4%
- Cache read$0.091/M$0.16/M▲ 75.4%
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via aki-io
What changed
- Cache read—$0.18/M
via aki-io
What changed
First observed in the model catalog
via openrouter
What changed
- structured outputYesNoRemoved
via opper
What changed
opper began listing this model
via chutes
What changed
- Input price$0.4/M$0.35/M▼ 12.5%
- Output price$3/M$2.75/M▼ 8.3%
- Cache read$0.04/M$0.035/M▼ 12.5%
via opper
What changed
opper began listing this model
via aki-io
What changed
First observed in the model catalog
via kilo
What changed
- Context205K197K−4%
via aki-io
What changed
Removed from the model catalog
via opper
What changed
opper began listing this model
via openrouter
What changed
- Output limit8K16K2×
- Input price$0.13/M$0.12/M▼ 7.7%
- Output price$0.52/M$0.5/M▼ 3.8%
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via openrouter
What changed
- Input price$0.3/M$0.24/M▼ 20.0%
- Output price$1.20/M$0.96/M▼ 20.0%
- Cache read$0.06/M$0.048/M▼ 20.0%
via openrouter
What changed
- Output limit66K262K4×
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via openrouter
What changed
- Input price$0.95/M$1/M▲ 5.3%
- Cache read$0.16/M$0.17/M▲ 6.3%
via merge-gateway
What changed
- Input price$0.22/M$0.035/M▼ 84.1%
- Output price$0.66/M$0.07/M▼ 89.4%
via kilo
What changed
- Output limit66K262K4×
via kilo
What changed
- Input price$0.064/M$0.06/M▼ 6.1%
- Output price$0.128/M$0.18/M▲ 40.8%
- Cache read$0.013/M$0.028/M▲ 119.1%
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via nano-gpt
What changed
First observed in the model catalog
via requesty
What changed
- Input price$4.50/M$3.60/M▼ 20.0%
- Output price$27/M$18/M▼ 33.3%
- Cache read$0.45/M$0.36/M▼ 20.0%
- cache write—$4.50/M
- +1 more changes
via opper
What changed
opper began listing this model
via kilo
What changed
- Input price$0.05/M$0.042/M▼ 16.0%
- Output price$0.25/M$0.22/M▼ 12.0%
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via kilo
What changed
- Context131K41K−69%
- Output limit8K16K2×
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via openrouter
What changed
- Input price$0.414/M$0.397/M▼ 4.1%
- Output price$0.828/M$0.794/M▼ 4.1%
- Cache read$0.034/M$0.033/M▼ 4.1%
via openrouter
What changed
- Input price$0.057/M$0.056/M▼ 2.4%
- Output price$0.115/M$0.112/M▼ 2.4%
- Cache read$0.011/M$0.011/M▼ 2.4%
via nano-gpt
What changed
Removed from the model catalog
via opper
What changed
opper began listing this model
via kilo
What changed
- structured outputYesNoRemoved
via edenai
What changed
- tool callNoYesEnabled
- structured outputNoYesEnabled
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via edenai
What changed
- Input price$0.176/M$0.352/M▲ 100.0%
- Output price$0.528/M$1.06/M▲ 100.0%
via opper
What changed
opper began listing this model
via nano-gpt
What changed
- tool callNoYesEnabled
- structured outputNoYesEnabled
- modalities.inputtext,image,audiotext,image,video,audio,pdf
- Context1.05M1.05M−0%
- +5 more changes
via opper
What changed
opper began listing this model
via opper
What changed
opper began listing this model
via openrouter
What changed
- Input price$0.064/M$0.06/M▼ 6.1%
- Output price$0.128/M$0.18/M▲ 40.8%
- Cache read$0.013/M$0.028/M▲ 119.1%
via kilo
What changed
- Context198K200K1×
- Output limit33K128K3.9×
via openrouter
What changed
- Input price$0.58/M$0.56/M▼ 3.3%
- Output price$2.44/M$2.36/M▼ 3.3%
- Cache read$0.098/M$0.094/M▼ 3.3%
via crof
What changed
Removed from the model catalog
via edenai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via runinfra
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via llmgateway
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit33K128K3.9×
- Output price$0.95/M$1.08/M▲ 13.7%
- Cache read$0.03/M$0.027/M▼ 10.0%
via edenai
What changed
- Context1.05M1M−5%
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via kilo
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via edenai
What changed
- Input price$0.603/M$1.16/M▲ 92.8%
- Output price$1.81/M$3.48/M▲ 92.8%
via kilo
What changed
First observed in the model catalog
via crof
What changed
Removed from the model catalog
via nano-gpt
What changed
First observed in the model catalog
via requesty
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via fireworks-ai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via vercel
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via crof
What changed
Removed from the model catalog
via openrouter
What changed
- Output limit16K262K16×
- Input price$0.09/M$0.1/M▲ 11.1%
- Cache read—$0.07/M
via vercel
What changed
- Input price$0.13/M$0.076/M▼ 41.5%
- Output price$0.26/M$0.153/M▼ 41.2%
- Cache read$0.028/M$0.014/M▼ 50.0%
via kilo
What changed
- Input price$0.14/M$0.44/M▲ 214.3%
- Output price$0.28/M$1.32/M▲ 371.4%
via nano-gpt
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via openrouter
What changed
Removed from the model catalog
via crossmodel
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit393K384K−2%
- Input price$1.60/M$0.549/M▼ 65.7%
- Output price$3.20/M$1.10/M▼ 65.7%
- Cache read$0.135/M$0.046/M▼ 66.1%
via crof
What changed
Removed from the model catalog
via kilo
What changed
First observed in the model catalog
via nano-gpt
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via edenai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via openrouter
What changed
- Output limit33K161K4.9×
- Input price$0.25/M$0.55/M▲ 120.0%
- Output price$0.95/M$1.65/M▲ 73.7%
- Cache read$0.13/M$0.55/M▲ 323.1%
via openrouter
What changed
First observed in the model catalog
via google-vertex
What changed
First observed in the model catalog
via llmgateway-providers
What changed
Removed from the model catalog
via hyper
What changed
- Output limit131K16K−88%
via openrouter
What changed
- Input price$0.45/M$0.4/M▼ 11.1%
- Output price$3.20/M$3/M▼ 6.3%
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via kilo
What changed
- Context1M1.05M1×
- Output limit262K131K−50%
via huggingface
What changed
- last updated2026-08-122026-08-22
via openrouter
What changed
- Input price$0.3/M$0.35/M▲ 16.7%
- Output price$1.10/M$1.50/M▲ 36.4%
via umans-ai-coding-plan
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via openrouter
What changed
- Input price$0.03/M$0.037/M▲ 23.3%
- Cache read$0.03/M—
via openrouter
What changed
- Output limit1.05M262K−75%
- Input price$0.065/M$0.064/M▼ 1.7%
- Output price$0.18/M$0.128/M▼ 29.0%
- Cache read$0.02/M$0.013/M▼ 36.1%
via openrouter
What changed
- Context256K131K−49%
- Input price$0.094/M$0.075/M▼ 20.0%
- Output price$0.25/M$0.2/M▼ 20.0%
via vercel
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via merge-gateway
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via merge-gateway
What changed
- attachmentYesNoRemoved
- modalities.inputtext,imagetext
- Context256K131K−49%
- Output limit64K33K−49%
via llmgateway
What changed
- Context1M1.05M1×
via kilo
What changed
- Context164K161K−2%
- Output limit33K161K4.9×
via openrouter
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via togetherai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via openrouter
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via llmgateway-providers
What changed
First observed in the model catalog
via nvidia
What changed
First observed in the model catalog
via deepseek
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via kilo
What changed
- Context1.05M262K−75%
- Output limit1.05M262K−75%
- Input price$0.065/M$0.064/M▼ 1.7%
- Output price$0.18/M$0.128/M▼ 29.0%
- +1 more changes
via openrouter
What changed
- Output limit262K131K−50%
via kilo
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via llmgateway-providers
What changed
Removed from the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Input price$0.22/M$0.44/M▲ 100.0%
- Output price$0.66/M$1.32/M▲ 100.0%
- Cache read$0.0070/M$0.014/M▲ 100.0%
via openrouter
What changed
- Output limit66K164K2.5×
- Input price$0.269/M$0.26/M▼ 3.3%
- Output price$0.4/M$0.38/M▼ 5.0%
- Cache read$0.135/M$0.13/M▼ 3.3%
via ofox
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via edenai
What changed
- Input price$0.14/M$0.22/M▲ 57.1%
- Output price$0.28/M$0.66/M▲ 135.7%
- Cache read$0.028/M$0.0070/M▼ 75.0%
via openrouter
What changed
- Input price$1.19/M$1.12/M▼ 5.6%
- Output price$3.56/M$3.37/M▼ 5.6%
- Cache read$0.04/M$0.037/M▼ 5.6%
via edenai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via openrouter
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via edenai
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
via baseten
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via edenai
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
via merge-gateway
What changed
- attachmentNoYesEnabled
- modalities.inputtexttext,image
- Context131K256K2×
- Output limit33K64K2×
via hyper
What changed
- open weightsNoYesEnabled
via arcee
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via digitalocean
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via edenai
What changed
- Context1.05M1M−5%
- Input price$0.66/M$1.32/M▲ 100.0%
- Output price$1.98/M$3.96/M▲ 100.0%
- Cache read$0.022/M$0.044/M▲ 100.0%
via vercel
What changed
- Context262K1M3.8×
- Output limit131K33K−75%
- Input price$0.05/M$0/M▼ 100.0%
- Output price$0.2/M$0/M▼ 100.0%
- +1 more changes
via kilo
What changed
- Context256K128K−50%
- Input price$0.094/M$0.075/M▼ 20.0%
- Output price$0.25/M$0.2/M▼ 20.0%
via openrouter
What changed
- Input price$0.081/M$0.078/M▼ 3.5%
- Output price$0.162/M$0.157/M▼ 3.5%
- Cache read$0.016/M$0.016/M▼ 3.5%
via merge-gateway
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$24/M▼ 20.0%
via alibaba-token-plan
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via kilo
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
via vercel
What changed
First observed in the model catalog
via umans-ai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via opencode
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via crof
What changed
Removed from the model catalog
via kilo
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.201/M$0.352/M▲ 75.2%
- Output price$0.603/M$1.06/M▲ 75.2%
via alibaba-token-plan-cn
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via crof
What changed
Removed from the model catalog
via llmgateway-providers
What changed
First observed in the model catalog
via kilo
What changed
- Input price$0.575/M$0.5/M▼ 13.0%
- Output price$3.45/M$3/M▼ 13.0%
- Cache read$0.115/M$0.1/M▼ 13.0%
- cache write$0.719/M$0.625/M▼ 13.0%
via deepinfra
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via google-vertex-anthropic
What changed
First observed in the model catalog
via deepseek
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via llmgateway
What changed
- Context1M1.05M1×
via kilo
What changed
First observed in the model catalog
via empiriolabs
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via edenai
What changed
First observed in the model catalog
via openrouter
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via openrouter
What changed
Removed from the model catalog
via digitalocean
What changed
- Context262K1M3.8×
via openrouter
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via kilo
What changed
- Output limit16K262K16×
via vercel
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via ofox
What changed
First observed in the model catalog
via cloudflare-workers-ai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via nvidia
What changed
First observed in the model catalog
via crof
What changed
First observed in the model catalog
via nano-gpt
What changed
- Output limit16K8K−50%
- Output price$0.4/M$1.80/M▲ 350.0%
- Cache read$0.2/M$0.4/M▲ 100.0%
via llmgateway-providers
What changed
First observed in the model catalog
via llmgateway
What changed
Removed from the model catalog
via kilo
What changed
- Input price$5/M$4/M▼ 20.0%
- Output price$30/M$20/M▼ 33.3%
- Cache read$0.5/M$0.4/M▼ 20.0%
- cache write$6.25/M$5/M▼ 20.0%
via kilo
What changed
Removed from the model catalog
via edenai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via kilo
What changed
- Input price$2.50/M$2/M▼ 20.0%
- Output price$15/M$10/M▼ 33.3%
- Cache read$0.25/M$0.2/M▼ 20.0%
- cache write$3.13/M$2.50/M▼ 20.0%
via kilo
What changed
- Output limit66K164K2.5×
via edenai
What changed
- open weightsNoYesEnabled
- last updated2026-08-122026-08-22
via kilo
What changed
- Context1.05M1.02M−2%
- Output limit393K384K−2%
via kilo
What changed
First observed in the model catalog
via vercel
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.402/M$0.201/M▼ 50.0%
- Output price$1.21/M$0.603/M▼ 50.0%
via scaleway
What changed
First observed in the model catalog
via nano-gpt
What changed
- Input price$0.13/M$0.08/M▼ 38.5%
- Output price$0.4/M$0.33/M▼ 17.5%
- Cache read$0.065/M$0.04/M▼ 38.5%
via hyper
What changed
- Input price$0.83/M$0.91/M▲ 9.6%
- Output price$2.56/M$2.93/M▲ 14.7%
- cache write$0.415/M$0.455/M▲ 9.6%
via edenai
What changed
- Input price$1.05/M$1.05/M▲ 0.2%
- Output price$1.05/M$1.05/M▲ 0.2%
via hyper
What changed
- Input price$0.55/M$0.544/M▼ 1.1%
- Output price$2.88/M$2.85/M▼ 1.0%
- cache write$0.275/M$0.272/M▼ 1.1%
via edenai
What changed
- Input price$0.467/M$0.468/M▲ 0.2%
- Output price$0.934/M$0.936/M▲ 0.2%
via opencode
What changed
- nameOx Alpha FreeOx Alpha Free (Unlimited)
via opencode-go
What changed
First observed in the model catalog
via ofox
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
via github-copilot
What changed
- Input price$5/M$2.50/M▼ 50.0%
- Output price$30/M$15/M▼ 50.0%
- Cache read$0.5/M$0.25/M▼ 50.0%
- cache write$6.25/M$3.13/M▼ 50.0%
via edenai
What changed
- Input price$0.175/M$0.175/M▲ 0.2%
- Output price$0.701/M$0.702/M▲ 0.2%
via openrouter
What changed
- Input price$0.083/M$0.081/M▼ 1.9%
- Output price$0.165/M$0.162/M▼ 1.9%
- Cache read$0.017/M$0.016/M▼ 1.9%
via kilo
What changed
- Output limit393K384K−2%
via opencode-go
What changed
First observed in the model catalog
via edenai
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via hyper
What changed
- Input price$0.12/M$0.11/M▼ 8.3%
- Output price$0.42/M$0.408/M▼ 2.9%
- cache write$0.06/M$0.055/M▼ 8.3%
via edenai
What changed
- Context8K262K32.8×
via edenai
What changed
- Input price$0.44/M$0.22/M▼ 50.0%
- Output price$1.32/M$0.66/M▼ 50.0%
- Cache read$0.014/M$0.0070/M▼ 50.0%
via openrouter
What changed
- Input price$0.95/M$0.58/M▼ 39.0%
- Output price$4/M$2.44/M▼ 39.0%
- Cache read$0.16/M$0.098/M▼ 39.0%
via venice
What changed
First observed in the model catalog
via kilo
What changed
- nameDeepSeek: DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp
- familydeepseekdeepseek-flash
via edenai
What changed
- Input price$1.32/M$0.66/M▼ 50.0%
- Output price$3.96/M$1.98/M▼ 50.0%
- Cache read$0.044/M$0.022/M▼ 50.0%
via edenai
What changed
- Input price$0.175/M$0.175/M▲ 0.2%
- Output price$0.759/M$0.76/M▲ 0.2%
via nano-gpt
What changed
- Input price$0.1/M$0.08/M▼ 20.0%
- Output price$0.35/M$0.33/M▼ 5.7%
- Cache read$0.05/M$0.04/M▼ 20.0%
via llmgateway-providers
What changed
First observed in the model catalog
via openrouter
What changed
- Output limit393K384K−2%
- Input price$0.14/M$0.08/M▼ 42.9%
- Output price$0.28/M$0.18/M▼ 35.7%
- Cache read$0.028/M$0.016/M▼ 42.9%
via kilo
What changed
- Input price$0.14/M$0.132/M▼ 5.7%
- Output price$0.58/M$0.528/M▼ 9.0%
- Cache read$0.035/M$0.033/M▼ 5.7%
via openrouter
What changed
- familydeepseekdeepseek-flash
via ofox
What changed
- Input price$1.50/M$0.75/M▼ 50.0%
- Output price$7.50/M$3.75/M▼ 50.0%
- Cache read$0.15/M$0.075/M▼ 50.0%
- cache write$0.083/M$0.042/M▼ 50.0%
via openrouter
What changed
First observed in the model catalog
via ofox
What changed
First observed in the model catalog
via edenai
What changed
- Input price$0.759/M$0.76/M▲ 0.2%
- Output price$0.759/M$0.76/M▲ 0.2%
via merge-gateway
What changed
First observed in the model catalog
via nano-gpt
What changed
First observed in the model catalog
via edenai
What changed
- Input price$1.21/M$0.603/M▼ 50.0%
- Output price$3.62/M$1.81/M▼ 50.0%
via kilo
What changed
First observed in the model catalog