The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
Specifications
| Provider | Meta |
|---|---|
| Context window | 131,072 tokens (131K) |
| Input price | $0.1 / 1M tokens |
| Output price | $0.32 / 1M tokens |
| Input modalities | text |
| Output modalities | text |
| Model ID | meta-llama/llama-3.3-70b-instruct |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does Llama 3.3 70B Instruct cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $0.1 per 1M input tokens and $0.32 per 1M output tokens. A provider's own list price may differ. On Vincony, Llama 3.3 70B Instruct is available on credit-based pricing alongside 800+ other models.
What is Llama 3.3 70B Instruct's context window?
Llama 3.3 70B Instruct supports a context window of up to 131,072 tokens (131K), which is how much text it can consider at once.
Who makes Llama 3.3 70B Instruct?
Llama 3.3 70B Instruct is built by Meta. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use Llama 3.3 70B Instruct?
You can use Llama 3.3 70B Instruct directly through Meta, or via Vincony, which aggregates 800+ models (including Llama 3.3 70B Instruct) into one interface with side-by-side comparison and credit-based pricing.