Capabilities and specifications
- Creator
- Meta
- Model family
- llama
- Knowledge cutoff
- 2023-12
- License
- Not documented
- Release date
- 2024-12-06
- Model ID
meta/llama-3.3-70b-instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
meta/llama-3.3-70b-instructProvider-specific identifiers, limits, and prices per million tokens.
| Provider | Provider model ID | Context | Input | Output | Cache read | Capabilities | Source |
|---|---|---|---|---|---|---|---|
meta/llama-3.3-70b |
128K | Not listed | Not listed | - | Tools | Docs | |
snowflake-llama3.3-70b |
128K | Not listed | Not listed | - | Tools | Docs | |
meta/llama-3.3-70b-instruct |
128K | Not listed | Not listed | - | ReasoningTools | Docs | |
meta/llama-3.3-70b-instruct |
128K | Not listed | Not listed | - | ToolsJSON | Docs | |
llama-3.3-70b-instruct |
128K | Not listed | Not listed | - | Tools | Docs | |
llama-3.3-70b-instruct |
131.072K | $0.13 | $0.4 | - | Tools | Docs | |
meta-llama/llama-3.3-70b-instruct |
131.072K | $0.13 | $0.4 | - | ToolsJSON | Docs | |
meta/llama-3.3-70b-instruct |
131.072K | $0.22 | $0.5 | $0.11 | Tools | Docs | |
meta-llama-3-3-70b-instruct |
128K | $0.5 | $1.5 | - | Tools | Docs | |
llama-3.3-70b-instruct |
128K | $0.513 | $1.045 | - | Tools | Docs | |
meta-llama/Llama-3.3-70B-Instruct |
131.072K | $0.59 | $0.79 | - | ToolsJSON | Docs | |
meta-llama/Meta-Llama-3.3-70B-Instruct |
131.072K | $0.59 | $0.79 | - | Tools | Docs | |
llama-3.3-70b-instruct |
128K | $0.71 | $0.71 | - | Tools | Docs | |
llama-3.3-70b-instruct |
128K | $0.71 | $0.71 | - | Tools | Docs | |
meta-llama/Llama-3.3-70B-Instruct |
131.072K | $0.9 | $0.9 | $0.9 | ReasoningTools | Docs | |
llama-3.3-70b-instruct |
100K | $0.9 | $0.9 | - | Tools | Docs | |
nvidia/Llama-3.3-70B-Instruct-FP8 |
128K | $1.15 | $1.15 | - | Tools | Docs | |
llama-3.3-70b-instruct |
100K | $1.254 | $1.254 | - | Tools | Docs | |
llama3-3-70b |
128K | $1.75 | $2.75 | - | Tools | Docs |
Capability badges appear only when the provider catalog explicitly lists support.
Observed serving measurements stay separate from model specifications and are never estimated from price or model size.
Generation speed; higher is faster.
Time until the first output token arrives.
Successful requests over a measured period.
Performance varies by inference provider, region, and load. Provider-level measurements appear only when a source reports them.
Raw results remain attached to their original source and version.
Detected changes from successful source imports.
Fetch the complete source-linked model record without an API key.
GET https://www.modelbench.lol/api/v1/models/meta/llama-3.3-70b-instructcurl "https://www.modelbench.lol/api/v1/models/meta/llama-3.3-70b-instruct"