Claims and evidence
- Status AvailableMeta publishes no deprecation page for Llama models; the weights of all four sizes and the quantized 1B and 3B builds are still listed in Meta's Hugging Face organisation (gated, license acceptance required) as seen on 2026-10-01. Primary source[4]
- Derived from (distillation) Llama 3.1The 1B and 3B models were pruned from Llama 3.1 8B and trained with logits from Llama 3.1 8B and 70B as token-level targets. The 11B and 90B Vision models start from pretrained Llama 3.1 text models and add a separately trained image adapter and encoder, leaving the language model weights unchanged. Primary source[1]Lightweight models; Vision models [2]Training Data [3]Model Architecture
- Successor of Llama 3.1 EditorialEditorial link along the Llama 3 version line: the announcement presents Llama 3.2 as the release after the Llama 3.1 herd, without calling it a successor. Primary source[1]
- Change · Modality Adds image input in the 11B and 90B Vision models; for image plus text tasks English is the only supported language.Compared with Llama 3.1 Primary source[1]Vision models [3]Model Information
- Change · Size Adds 1B and 3B text models for mobile and edge devices, enabled for Qualcomm and MediaTek hardware and optimised for Arm processors.Compared with Llama 3.1 Primary source[1]Takeaways; Lightweight models
- Change · Efficiency Quantized builds of the 1B and 3B Instruct models followed on 24 October 2024, limited to 8K context for mobile devices.Compared with Llama 3.1 Primary source[6]
- Change · Licensing The acceptable use policy does not grant the license rights for the multimodal models to individuals or companies based in the European Union.Compared with Llama 3.1 Primary source[7]paragraph on multimodal models after the list of prohibited uses
- Input textApplies to the 1B and 3B text models; the two Vision variants also accept images. Primary source[2]Model Information table, Input modalities
- Output text, code Primary source[2]Model Information table, Output modalities
- Open weights YesWeights under the custom Llama 3.2 Community License, after accepting the license. The acceptable use policy does not grant the license rights for the multimodal models to individuals or companies based in the European Union; end users of products that include them are exempt. Primary source[1]Takeaways [4] [7]paragraph on multimodal models after the list of prohibited uses
- Context window 128K tokensThe quantized 1B and 3B builds of October 2024 are limited to 8k. Primary source[1]Takeaways (1B and 3B) [2]Model Information table, Context Length (128k) [3]Model Information table, Context length (128k)
- Access Open-weights download, cloud partner, On device, consumer appOn-device: the 1B and 3B models (and their quantized builds) target mobile and edge hardware. Consumer access: Meta says the Vision models can be tried in Meta AI. Primary source[1]Takeaways [6]Takeaways
Lineage
Predecessors
- Llama 3.1 · 23 July 2024
Successors
- Llama 3.3 · 6 December 2024
Based on
- Llama 3.1 · 23 July 2024 · derived (distillation)
Variants and derived
None recorded.
Siblings
None recorded.
All ancestors
- LLaMA · 24 February 2023
- Llama 2 · 18 July 2023
- Meta Llama 3 · 18 April 2024
- Llama 3.1 · 23 July 2024
All descendants
- Llama 3.3 · 6 December 2024
- Sonar (February 2025) · Perplexity · 11 February 2025
- Llama-3.3-Nemotron-Super-49B-v1 · NVIDIA · 18 March 2025
- Llama 4 Maverick · 5 April 2025
- Llama 4 Scout · 5 April 2025
- Llama-3.3-Nemotron-Super-49B-v1.5 · NVIDIA · 25 July 2025
- Nemotron 3 Super · NVIDIA · 11 March 2026
- Muse Spark · 8 April 2026
- Muse Spark 1.1 · 9 July 2026
- Muse Spark 1.2 · 5 August 2026
- Muse Glimmer · 10 August 2026
- Muse Spark 1.3 · 2 September 2026
Variants
Llama 3.2 1B
Same dates as Llama 3.2.
- Variant Inline variant in this record.Text-only, pretrained and instruction-tuned. Quantized QLoRA and SpinQuant builds of the Instruct model followed on 2024-10-24, with context limited to 8k. Primary source[1]Takeaways; Lightweight models [2]Model Information [6] [4]
Also known as: Llama-3.2-1B, Llama-3.2-1B-Instruct, Llama-3.2-1B-Instruct-QLORA_INT4_EO8, Llama-3.2-1B-Instruct-SpinQuant_INT4_EO8
Differs in:
- Parameters: 1B (1.23B) Evidence not assessed [2]Model Information table, Params
Llama 3.2 3B
Same dates as Llama 3.2.
- Variant Inline variant in this record.Text-only, pretrained and instruction-tuned. Quantized QLoRA and SpinQuant builds of the Instruct model followed on 2024-10-24, with context limited to 8k. Primary source[1]Takeaways; Lightweight models [2]Model Information [6] [4]
Also known as: Llama-3.2-3B, Llama-3.2-3B-Instruct, Llama-3.2-3B-Instruct-QLORA_INT4_EO8, Llama-3.2-3B-Instruct-SpinQuant_INT4_EO8
Differs in:
- Parameters: 3B (3.21B) Evidence not assessed [2]Model Information table, Params
Llama 3.2 11B Vision
Same dates as Llama 3.2.
- Variant Inline variant in this record.Pretrained and instruction-tuned. For image plus text tasks the model card lists English as the only supported language. Primary source[1]Takeaways; Vision models [3]Model Information [4]
Also known as: Llama-3.2-11B-Vision, Llama-3.2-11B-Vision-Instruct, Llama 3.2 11B
Differs in:
Llama 3.2 90B Vision
Same dates as Llama 3.2.
- Variant Inline variant in this record.Pretrained and instruction-tuned. For image plus text tasks the model card lists English as the only supported language. Primary source[1]Takeaways; Vision models [3]Model Information [4]
Also known as: Llama-3.2-90B-Vision, Llama-3.2-90B-Vision-Instruct, Llama 3.2 90B
Differs in:
Related AI Radar coverage
AI Radar coverage starts in June 2026; no coverage linked yet.