Llama is a family of large language models developed by Meta and released, in large part, with open weights for research and commercial use. First introduced in 2023 with LLaMA, followed by Llama 2, Llama 3 and later versions, the models are transformer-based and trained on large text corpora. By releasing model weights under permissive terms, Meta enabled a wide range of independent fine-tuning and deployment, making Llama a common base for open models. The family spans several parameter sizes to suit different compute and latency requirements.

Semantic Classification

Content

  • Llama is a series of decoder-only transformer language models from Meta. The initial LLaMA release in early 2023 demonstrated that comparatively smaller models trained on more data could match or exceed larger models on many benchmarks, which influenced subsequent training practice across the field.
  • A defining characteristic of the family is the release of model weights, initially under research terms and later under licences permitting broad commercial use. This availability seeded a large community of fine-tuned derivatives and tooling, and made Llama a default starting point for organisations building on open models.
  • Successive versions improved scale, training data, context length and instruction-following, and added multimodal and larger variants. The models are released in multiple sizes so that practitioners can trade off quality against the compute and memory needed for training and inference.

Provenance