Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
Specifications
| Provider | Meta |
|---|---|
| Context window | 131,072 tokens (131K) |
| Input price | $0.05 / 1M tokens |
| Output price | $0.08 / 1M tokens |
| Input modalities | text |
| Output modalities | text |
| Model ID | meta-llama/llama-3.1-8b-instruct |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does Llama 3.1 8B Instruct cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $0.05 per 1M input tokens and $0.08 per 1M output tokens. A provider's own list price may differ. On Vincony, Llama 3.1 8B Instruct is available on credit-based pricing alongside 800+ other models.
What is Llama 3.1 8B Instruct's context window?
Llama 3.1 8B Instruct supports a context window of up to 131,072 tokens (131K), which is how much text it can consider at once.
Who makes Llama 3.1 8B Instruct?
Llama 3.1 8B Instruct is built by Meta. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use Llama 3.1 8B Instruct?
You can use Llama 3.1 8B Instruct directly through Meta, or via Vincony, which aggregates 800+ models (including Llama 3.1 8B Instruct) into one interface with side-by-side comparison and credit-based pricing.