Llama-3.3-70B-Instruct
Summary
Llama-3.3-70B-Instruct is an open-weight language model from Meta, released on 6 Dec 2024. It accepts text and generates text, with a 128K-token context window and up to 4K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $0.14 per million input tokens and $0.40 per million output tokens (Novita AI), across 8 providers we track.
Open Llama instruction model for multilingual chat, reasoning, and coding
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | Meta |
|---|---|
| Release date | 6 Dec 2024 [models.dev] |
| Status | Generally available [models.dev] |
| Weights | Open weights [Hugging Face] |
| License | llama3.3 [Hugging Face] |
| Parameters | 70.6B [Hugging Face] |
| Context window | 128K (128,000 tokens) [models.dev] |
| Max output | 4K tokens [models.dev] |
| Knowledge cutoff | Dec 2023 [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | No [models.dev] |
| Tool calling | Yes [models.dev] |
| Hugging Face | meta-llama/Llama-3.3-70B-Instruct [Hugging Face] |
Llama-3.3-70B-Instruct API pricing by provider
USD per million tokens. Sorted by input price.
| Provider | Input | Output | Cached input | Context | Source |
|---|---|---|---|---|---|
| Llama API llama-3.3-70b-instruct | Free | Free | – | 128K | models.dev, 4 Oct 2026 |
| Novita AI meta-llama/llama-3.3-70b-instruct | $0.14 | $0.40 | – | 131K | models.dev, 4 Oct 2026 |
| OpenRouter meta-llama/llama-3.3-70b-instruct | $0.22 | $0.50 | $0.11 | 131K | OpenRouter, 4 Oct 2026 |
| Cloudflare Workers AI @cf/meta/llama-3.3-70b-instruct-fp8-fast | $0.29 | $2.25 | – | 24K | models.dev, 4 Oct 2026 |
| Hugging Face Inference Providers meta-llama/Llama-3.3-70B-Instruct | $0.59 | $0.79 | – | 131K | models.dev, 4 Oct 2026 |
| Azure AI Foundry llama-3.3-70b-instruct | $0.71 | $0.71 | – | 128K | models.dev, 4 Oct 2026 |
| Amazon Bedrock meta.llama3-3-70b-instruct-v1:0 | $0.72 | $0.72 | – | 128K | models.dev, 4 Oct 2026 |
| Amazon Bedrock us.meta.llama3-3-70b-instruct-v1:0 | $0.72 | $0.72 | – | 128K | models.dev, 4 Oct 2026 |
| Snowflake Cortex snowflake-llama3.3-70b | – | – | – | 128K | models.dev, 4 Oct 2026 |
Frequently asked questions
What is the context window of Llama-3.3-70B-Instruct?
Llama-3.3-70B-Instruct has a context window of 128,000 tokens (128K) and can generate up to 4,096 tokens in a single response, according to models.dev.
How much does the Llama-3.3-70B-Instruct API cost?
Through Llama API, Llama-3.3-70B-Instruct costs Free per million input tokens and Free per million output tokens. Prices last verified 4 Oct 2026.
Is Llama-3.3-70B-Instruct open source?
Llama-3.3-70B-Instruct is an open-weight model: its weights are published on Hugging Face as meta-llama/Llama-3.3-70B-Instruct under the llama3.3 license. Check the license terms before commercial use.
When was Llama-3.3-70B-Instruct released?
Meta released Llama-3.3-70B-Instruct on 6 Dec 2024, according to models.dev.
Which providers offer Llama-3.3-70B-Instruct?
We track Llama-3.3-70B-Instruct on 8 providers: Llama API, Novita AI, OpenRouter, Cloudflare Workers AI, Hugging Face Inference Providers, Azure AI Foundry, Amazon Bedrock, Snowflake Cortex.
Change history
No changes detected since we started tracking this model. We re-check every source daily.
Sources
- https://openrouter.ai/meta-llama/llama-3.3-70b-instruct
- https://github.com/sst/models.dev/blob/dev/providers/llama/models/llama-3.3-70b-instruct.toml
- https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct
- https://github.com/sst/models.dev/blob/dev/providers/novita-ai/models/meta-llama/llama-3.3-70b-instruct.toml
- https://github.com/sst/models.dev/blob/dev/providers/cloudflare-workers-ai/models/@cf/meta/llama-3.3-70b-instruct-fp8-fast.toml
- https://github.com/sst/models.dev/blob/dev/providers/huggingface/models/meta-llama/Llama-3.3-70B-Instruct.toml
- https://github.com/sst/models.dev/blob/dev/providers/azure/models/llama-3.3-70b-instruct.toml
- https://github.com/sst/models.dev/blob/dev/providers/amazon-bedrock/models/meta.llama3-3-70b-instruct-v1:0.toml
- https://github.com/sst/models.dev/blob/dev/providers/amazon-bedrock/models/us.meta.llama3-3-70b-instruct-v1:0.toml
- https://github.com/sst/models.dev/blob/dev/providers/snowflake-cortex/models/snowflake-llama3.3-70b.toml