Skip to main content

Meta · Text

Meta Llama models: Versions, specs, GPU requirements & hosting

Meta Llama provides widely adopted open-weight generations; Scout and Maverick are sibling Llama 4 variants rather than sequential generations.

Technical deployment reference illustration for Meta Llama
Reference deployment flow; exact architecture facts are listed separately.

Overview and history

Meta Llama is an official Meta model family represented here through independently verified release, checkpoint, license, runtime and hardware evidence. Meta Llama provides widely adopted open-weight generations; Scout and Maverick are sibling Llama 4 variants rather than sequential generations. Current public releases are grouped by the General tracks so sibling sizes and specialist branches are not presented as false chronological generations.

FusionPointAI currently tracks 3 certified public Meta Llama releases. The catalogue preserves release dates and evaluates Current, Recent, Previous and Legacy status within each model track rather than by database insertion order.

Broad ecosystem
Multimodal Llama 4 variants
Mature deployment tooling

Architecture and model range

The verified public range spans 109B to 405B total parameters and context windows up to 10,000,000 tokens. Each release states whether it is dense or mixture-of-experts and records active parameters separately when the official source publishes them.

Public releases3
Model tracks1
Current generationLlama 4 Scout 17B-16E
Last verified2026-08-26T00:00:00.000000Z

Deployment ecosystem

Choose a specific Meta Llama release and certified checkpoint before sizing hardware. FusionPointAI rejects unknown runtime configurations, calculates resident memory from total model weights, then distinguishes globally compatible GPUs from currently rentable inventory.

Validate the exact checkpoint license, runtime version, precision kernel, context length, concurrency and multi-GPU topology. Active MoE parameters affect compute, but resident model memory continues to use total stored weights.

Common use cases

General assistants
Private fine-tuning
Multimodal applications

License overview

Licenses are certified per release, not inferred from the provider. Review the release page before commercial deployment because terms can differ between generations and variants.

Commercial-use status is a source-backed catalogue classification, not legal advice. Follow the linked official license and attribution requirements.

Release timeline

meta/llama/llama-4-scout-17b-16e @ 15fb562bcf144f0415ac8dbc9c7aee1e174d2e826178c9691256acdb94156442

Official source

meta/llama/llama-4-maverick-17b-128e @ 0b389dbe619938c2b9ee945f693f2816c90ed1c6f2754c113b652873e163735f

Official source

meta/llama/llama-3-1-405b @ ec5a654d6f9ae0f8c3e0ad5cf9d102d9937c302c51a497a7ae316cdb15dfed83

Official source

Frequently asked questions

Which Meta Llama release is current?

Current releases are Llama 4 Scout 17B-16E, Llama 4 Maverick 17B-128E, evaluated within their respective tracks.

Can Meta Llama be self-hosted?

Only releases with a verified official checkpoint and supported runtime configuration are presented as deployable.

Which GPU should run Meta Llama?

GPU selection depends on the exact variant, checkpoint precision, context length and concurrency; use the release hardware recommendations rather than the family name alone.

Official references

Last verified: 2026-08-26