Llama 5

Meta's next flagship open-weight multimodal model, built for on-device and datacenter use alike

Free Multimodal
Visit Product Page →
2000000 context tokens
520B (24B active) parameters
Llama 4 Community License license
Jul 2026 released

Llama 5 is Meta’s follow-up to last year’s Llama 4 line, again shipped as a mixture-of-experts model but scaled up to 520 billion total parameters with 24 billion active per token. The bigger jump is the 2 million token context window, double what Llama 4 Maverick offered, which Meta is pitching for tasks like ingesting entire codebases or long video transcripts in one pass. It handles text, image, and audio input natively rather than bolting encoders onto a text-only base, and Meta says it edges out Gemini 3.5 Pro and GPT-5.6 on several open benchmarks while still running efficiently thanks to the sparse expert routing.

As with prior Llama releases, the weights are open under the Llama 4 Community License and available immediately on Hugging Face, with hosting partners like Together AI, AWS Bedrock, and Azure AI Foundry offering managed inference for teams that don’t want to run it themselves. Meta continues to withhold full commercial rights from companies above a certain user threshold, the same carve-out that’s applied since Llama 2.