Milestones
- LLaMA released to researchers24 Feb 2023Complete.
- Llama 2 with commercial licence18 Jul 2023Complete.
- Llama 318 Apr 2024Complete.
- Llama 3.1 405B, largest open model at the time23 Jul 2024Complete.
- Llama 3.2 vision and on-device models25 Sep 2024Complete.
- Llama 4 Scout and Maverick (mixture-of-experts)5 Apr 2025Complete.
- Llama 4 Behemoth releaseMissedMissed.
Most important updates
- 5 Apr 2025
- 23 Jul 2024
- 18 Apr 2024
- 18 Jul 2023
- 24 Feb 2023
Current obstacles
- Chinese open models caught upDeepSeek, Qwen and Kimi releases in 2025 overtook Llama on many open-model rankings.
- Benchmark trustThe Llama 4 version that was tested differed from the one released, raising doubts about the scores.
Physics limits
- Shrinking numbers loses informationRunning on smaller hardware means storing weights in 4 to 8 bits instead of 16. Below about 4 bits accuracy drops quickly, so model size sets a floor on memory needed.
- Running out of human-written textFrontier models already train on 15 to 40 trillion tokens. Estimates put the usable stock of public human text at a few hundred trillion, so data, not chips, starts to cap plain scaling.
- Trained for likely text, not true textThe model learns to produce plausible words. Facts seen only once or twice in training can't be recalled reliably, so some rate of confident errors is built into the method.
How it works

Download and run
Anyone can run Llama on their own hardware, retrain it on private data or shrink it to fit a laptop or phone.
Mixture of experts
Llama 4 Maverick has about 400B parameters but uses only 17B per word, sending each word to the right experts.
Images from the start
Pictures and text enter the same model from the first layer, instead of through a separate vision add-on.
Papers & demos
- Jul 2024paperThe Llama 3 Herd of ModelsExplained how a 405B dense model was trained on about 15 trillion tokens using 16,000 H100 chips.
- Feb 2023paperLLaMA: Open and Efficient Foundation Language ModelsShowed smaller models trained on more data can match bigger ones. It started the open-weight wave.
Update log
Sat 5 Apr 2025
- Major: BlogSoftware
Wed 25 Sep 2024
- Minor: BlogSoftware
Tue 23 Jul 2024
- Major: BlogSoftware
Thu 18 Apr 2024
- Major: BlogSoftware
Tue 18 Jul 2023
- Major: BlogSoftware
Fri 24 Feb 2023
- Major: BlogSoftware
About Meta
Meta
Llama open models, Muse, AI data centers
Meta, owner of Facebook, Instagram and WhatsApp, made open-weight AI mainstream with its Llama models from 2023. In 2025 it formed Meta Superintelligence Labs. In April 2026 it released Muse Spark, its first closed frontier model.
- Llama was downloaded over a billion times, but Llama 4 (April 2025) was seen as behind the leaders.
- Paid about $14.3B for 49% of Scale AI and hired its CEO Alexandr Wang to lead its new AI lab (June 2025).
- Muse Spark (April 2026) is closed, a shift away from open weights.
- Founded
- 200422 yrs
- Headquarters
- United States
- Status
- Public
- Listed
- META (opens google.com)NASDAQ
- Valuation
- public
- Staff
- ~74,067Dec 2024
- Coverage
- 3 programs · 19 updateslatest 17d agochecked 25 Sep
- Partners
- Scale AIBlue Owl CapitalMicrosoft
- People
- Mark ZuckerbergFounder, Chairman and CEOAlexandr WangChief AI Officer

