Llama 3.3 70B Instruct
Meta
meta-llama/llama-3.3-70b-instructLlama 3.3 70B Instruct is Meta's open-weight model that delivers performance close to Llama 3.1 405B at a fraction of the size and cost, with a 128K context window. It is well suited for general chat, reasoning, and coding in self-hosted deployments.
Cost rate
0.1x
Context
131K
Released
Dec 6, 2024
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
Legal & Government
#217 · top 59%
Medicine & Healthcare
#225 · top 62%
Life, Physical, & Social Science
#231 · top 59%
Mathematical
#233 · top 63%
Performance
Median latency and throughput measured across recent requests.
Throughput88 tok/s
Latency