Llama-3.3-70B-Instruct
About Llama-3.3-70B-Instruct
Llama 3.3 is the most advanced multilingual open-source large language model in the Llama series, offering performance comparable to a 405B model at a significantly lower cost. Built on the Transformer architecture, it enhances usefulness and safety through supervised fine-tuning (SFT) and reinforcement learning from human feedback (RLHF). Its instruction-tuned version is optimized for multilingual dialogue and outperforms many open-source and closed chat models across various industry benchmarks. The knowledge cutoff is December 2023.
Metadata
Specification
State
Deprecated
Architecture
Calibrated
No
Mixture of Experts
No
Total Parameters
70B
Activated Parameters
Reasoning
No
Precision
FP8
Context length
33K
Max Tokens
Compare with Other Models
See how this model stacks up against others.

