Llama is Meta’s open-weight language model family. Since the original 2023 release it has become the most widely fine-tuned foundation model in the world, with variants spanning 1B to 405B parameters across dense and Mixture-of-Experts architectures.
See all models from MetaModels in family
13
Open weight
13
API only
0
Avg score
39.2
Top benchmark
88.9
MATH-500
Total HF downloads
11.9M
Primary modality
Text
Context window
2K – 10.0M
First release
Feb 2023
Latest release
Oct 2025
Every release in the Llama family, ranked by composite score across benchmarks, popularity, efficiency, and versatility.
| # | Model | Modality | Score | Params | Released |
|---|---|---|---|---|---|
| 1 | embedding | BB62.5 | 7.5B | Oct 2025 | |
| 2 | text | CC54.8 | 8B | Jul 2024 | |
| 3 | text | CC50.1 | 8B | Apr 2024 | |
| 4 | text | CC45.2 | 70B | Dec 2024 | |
| 5 | text | CC40.7 | 70B | Jul 2024 | |
| 6 | text | CC40.7 | 109B | Apr 2025 | |
| 7 | text | DD39.8 | 400B | Apr 2025 | |
| 8 | text | DD38.9 | 7B | Jul 2023 | |
| 9 | text | DD35.4 | 13B | Jul 2023 | |
| 10 | text | DD33.5 | 70B | Apr 2024 |
When a family spans multiple modalities, the strongest model in each one.
How the family splits across text, image, video, audio, and embedding releases.
When each release shipped, newest first. Useful for tracking version cadence.
Oct 21
Apr 4
Apr 4
Jul 22
Jul 22
Jul 22
Apr 17
Apr 17
Jul 17
Jul 17
Jul 17
Feb 23
Composite grades across this family. Higher is better, blending benchmarks, popularity, and efficiency.
Models with downloadable weights, ranked by composite score.
| # | Model | Modality | Score | Params | Released |
|---|---|---|---|---|---|
| 1 | embedding | BB62.5 | 7.5B | Oct 2025 | |
| 2 | text | CC54.8 | 8B | Jul 2024 | |
| 3 | text | CC50.1 | 8B | Apr 2024 | |
| 4 | text | CC45.2 | 70B | Dec 2024 | |
| 5 | text | CC40.7 | 70B | Jul 2024 | |
| 6 | text | CC40.7 | 109B | Apr 2025 | |
| 7 | text | DD39.8 | 400B | Apr 2025 | |
| 8 | text | DD38.9 | 7B | Jul 2023 | |
| 9 | text | DD35.4 | 13B | Jul 2023 | |
| 10 | text | DD33.5 | 70B | Apr 2024 | |
| 11 | text | DD33.2 | 405B | Jul 2024 | |
| 12 | text | DD27.2 | 70B | Jul 2023 | |
| 13 | text | FF7.3 | 65B | Feb 2023 |
The Llama family is a series of AI models from Meta. This page lists every release in the family with its benchmark scores, parameter count, and hardware requirements.
By composite score, llama-embed-nemotron-8b is currently the top model in the family. For local inference, match the parameter count to your VRAM budget. For quality, pick the highest scorer that fits.
See the open-weight section above for models you can run locally. The API-only section lists closed releases that must be accessed through the provider’s API.
Spin up an instance in the cloud, or pick local hardware that fits.

Full Directory
Open the full directory to filter by hardware, capability, license, and benchmark score.

Or Browse by Provider
See every model from a lab side by side, with aggregate stats. Useful when you want a cross-family view of one provider.