Model
Llama-3.1-8B-Instruct
Compact open Llama model for lightweight chat, drafting, and self-hosting
Compare providers (sorted by listed input price)
15 of 17 providers publish a price for Llama-3.1-8B-Instruct. The cheapest full-context option is Kilo Gateway at $0.02/$0.04 per M. Listed prices span 9.0× for identical weights.
Best price history
Daily lowest published price across live providers. A lower line is better for the buyer.
- Best input /M
- $0.02unchanged
- Best output /M
- $0.04unchanged
- Provider coverage
- 18
- 17 at the start · 24 snapshots
Published best prices are unchanged.
24 daily snapshots collected from Aug 21, 2026 through Sep 14, 2026. A price chart will appear after the first real move.
Price badge
https://www.vaanalytics.in/badge/meta/llama-3.1-8b-instruct.svgLive SVG, regenerated on every hourly sync — paste it into a README to always show the current cheapest listed price. Updates when prices change; no build step needed on your side.
Recent changes
Changelog →- capabilityvia nano-gpt · Sep 13, 2026
- tool_call: false → true
- capabilityvia amazon-bedrock · Sep 13, 2026
- description: Open Llama instruction model for multilingual chat, reasoning, and coding → Compact open Llama model for lightweight chat, drafting, and self-hosting
- new modelvia amazon-bedrock · Sep 10, 2026
- contextvia wandb · Aug 27, 2026
- context: 128K → 131K tokens
- output: 128K → 131K tokens
- contextvia kilo · Aug 25, 2026
- output: 131K → 118K tokens
- contextvia openrouter · Aug 25, 2026
- output: 131K → 118K tokens
Weights
Commercial use allowed, but the lab attaches conditions — read them before shipping. Hugging Face requires you to accept terms, and the lab may approve access manually.
Derived from meta-llama/Meta-Llama-3.1-8B.
licence and gating from the Hugging Face model card ↗ — the card is authoritative, this is a summary
Prices are per million tokens (USD). “—” means the provider does not publicly list a price for this model. Data from models.dev; verify with the provider before purchasing.