Meta
Search
Documentation

Products
Muse Code
Overview
Meta Model API
Overview
API Login
Models
Muse
Muse Spark 1.3
Muse Glimmer
Muse Image
Muse Voice Transcribe
Llama
Llama 4
Llama 3
Resources
Documentation
Model API docs
Learn
Cookbooks
Videos
Blog
Case studies
Community
Github
Meta Models
Llama
Hugging Face
Meta Models
Safety
Llama Protections
Overview
Llama Defenders Program
Developer use guide

ModelsLlamaLlama 3
Stay updated
Get started
Meta
post image
post image
post image
post image
Products
Muse Code
Meta Model API
Models
Muse Spark 1.2
Muse Spark 1.1
Muse Glimmer
Muse Voice Transcribe
Llama 4
Llama 3
Documentation
Meta Model API Docs
Muse Glimmer Docs
Llama Docs
Resources
Cookbook
Blog
Videos
Case studies
FAQs
Community
Meta-Models Github
Llama GitHub
Hugging Face
Terms & policies
Terms of Service
Privacy Policy
Cookie Policy
Products
Muse Code
Meta Model API
Models
Muse Spark 1.3
Muse Spark 1.2
Muse Spark 1.1
Muse Glimmer
Muse Image
Muse Voice Transcribe
Llama 4
Llama 3
Documentation
Meta Model API Docs
Muse Glimmer Docs
Llama Docs
Resources
Cookbook
Blog
Videos
Case studies
FAQs
Community
Meta-Models Github
Llama GitHub
Hugging Face
Terms & policies
Terms of Service
Privacy Policy
Cookie Policy

Llama 3

The open-source AI models you can fine-tune, distill and deploy anywhere. Choose from our collection of models: Llama 3.1, Llama 3.2, Llama 3.3.
Download models

Llama 3 models

Llama includes multilingual text-only models (1B, 3B), including quantized versions, text-image models (11B, 90B) and Llama 3.3 70B model offering similar performance to the Llama 3.1 405B model, allowing developers to achieve greater quality and performance on text-based applications at a fraction of the cost.
MOST RECENT

Llama 3.3

View documentation
Llama 3.3 is a text-only 70B instruction-tuned model that provides enhanced performance relative to Llama 3.1 70B–and relative to Llama 3.2 90B when used for text-only applications. Moreover, for some applications, Llama 3.3 70B approaches the performance of Llama 3.1 405B.
VARIATIONS
70BLeading performance at a fraction of the cost.
Download
OLDER MODELS

Llama 3.2

View documentation
Llama 3.2 11B and 90B Vision-enabled models plus lightweight 1B and 3B models optimized for on-device and edge deployment.
VARIATIONS
1B & 3BLight-weight, efficient models you can run everywhere.
Download
11B & 90BMultimodal and can reason on high resolution images.
Download

Llama 3.1

View documentation
The open source AI model you can fine-tune, distill and deploy anywhere. Our instruction-tuned model is available in 8B, 70B and 405B versions.
VARIATIONS
8BLight-weight, ultra-fast model you can run anywhere.
Download
405BFlagship foundation model driving widest variety of use cases.
Download

Do more with Llama

llama 3 graphic
On-device
Use our 1B or 3B models for on device applications such as summarizing a discussion from your phone or calling on-device tools like calendar.
Learn more
Multimodal
Use our 11B or 90B models for image use cases such as transforming an existing image into something new or getting more information from an image of your surroundings.
Learn more
Llama Stack
Seamlessly build agentic applications from a comprehensive toolchain.
Learn more
View all capabilities

Start building with Llama 3

Download models
View documentation
SAFETY

Protections in the era of generative AI.

Comprehensive system-level protections proactively identify and mitigate potential risks, empowering developers to more easily deploy generative AI responsibly.
Protection tools accessible to everyone
Learn more
Enabling AI Defenders
Learn more

Llama 3 tuned benchmarks

Llama 3.3 instruction
Lightweight instruction
Vision instruction
Category
Benchmark

General

MMLU Chat

(0-shot, CoT)

MMLU PRO

(5-shot, CoT)

Instruction Following

IFEval

Code

HumanEval

(0-shot)

MBPP EvalPlus

(base) (0-shot)

Math

MATH

(0-sho, CoT)

Reasoning

GPQA Diamond

(0-shot, CoT)

Tool use

BFCL v2

(0-shot)

Long context

NIH/Multi-needle

Multilingual

Multilingual MGSM

(0-shot)

Pricing*

1M Input tokens

(Cheapest among providers)*

1M Output tokens

(Cheapest among providers)*

Llama 3.1 70B

86.0

66.4

87.5

80.5

86.0

67.8

48.0

77.5

97.5

86.9

$0.1

$0.4

Llama 3.3 70B

86.0

68.9

92.1

88.4

87.6

77.0

50.5

77.3

97.5

91.1

$0.1

$0.4

Amazon Nova
Pro

85.9

-

92.1

89.0

-

76.6

-

-

-

-

$0.80

$3.20

Llama 3.1 405B

88.6

73.4

88.6

89.0

88.6

73.9

49.0

81.1

98.1

91.6

$1.0

$1.8

Gemini Pro
1.5

87.1

76.1

81.9

89.0

87.8

82.9

53.5

80.3

94.7

89.6

$1.30

$5.0

GPT-4o

87.5

73.8

84.6

86.0

83.9

76.9

47.5

74.0

-

90.6

2.5$

10.0$

Claude 3.5
Sonnet

88.9

77.8

89.3

93.7

86.8

78.3

65.0

79.3

99.4

92.8

$3.0

$15.0

* API Pricing based on publicly available data on Artificial Analysis as of 12/3/24.