Skip to content
Lanseringar· NewsAvailable

Meta Launches Llama 4: MoE Architecture and Record-Breaking Context Window

Meta has launched the Llama 4 family featuring MoE architecture and native multimodality. The models support context windows of up to 10 million tokens and demonstrate high performance on local Apple hardware.

By the Aheadline editorial team·2 aug. 2026·2 min read·Source: Entity-watch: Meta AIVerifierad signalAI-generated
Meta Launches Llama 4: MoE Architecture and Record-Breaking Context Window
Meta Launches Llama 4: MoE Architecture and Record-Breaking Context Window
Meta Launches Llama 4: MoE Architecture and Record-Breaking Context Window
By · Policy- & EU-reporter
Last updated

What happened?

Meta has launched the Llama 4 model family, marking the company’s transition to MoE (Mixture of Experts) architecture and native multimodality. The family consists of three models: Llama 4 Scout (17B active parameters, 16 experts, 109B total), Llama 4 Maverick (17B active parameters, 128 experts, 402B total), and Behemoth (288B active parameters, 16 experts, 2T total). The models support a context window of up to 10 million tokens.

Key facts

Lanseringsdatum7 april 2025
Maximalt kontextfönster10 miljoner tokens
Största modellens parametrar2 000 miljarder (2T) totalt
MolnstödDatabricks Foundation Model API

Why it matters

The launch represents a significant leap for open-source AI by combining an extremely long context window with MoE architecture. By activating only a fraction of the parameters during each inference run, computational requirements are reduced, making it possible to run extremely large models on consumer hardware and local infrastructure.

Who is affected?

Developers, AI researchers, and companies wishing to run large-scale open-source models locally are directly affected. The efficient MoE architecture enables the execution of very large models on local hardware, such as Mac computers with Apple Silicon and high unified memory capacity.

Impact on the EU

The Llama 4 models are released as open source and are available globally, including within the EU. However, full-scale deployment in commercial services within the EU may be subject to future compliance requirements under the EU AI Act.

What else you should know

Initial tests showed that Llama 4 Maverick achieved a generation speed of 50 tokens per second on a single Mac equipped with an M3 Ultra chip. Conversely, independent tests showed varying results regarding code generation, with some users reporting shortcomings in complex coding tasks compared to specialized models.

Frequently asked questions

Quick answers about this story

Vad har hänt?
Meta har lanserat sin nya AI-modellfamilj Llama 4 med MoE-arkitektur och stöd för multimodala indata.
När hände det?
Meta lanserade Llama 4-modellerna i april 2025.
Varför spelar det roll?
Modellerna introducerar rekordstora kontextfönster på upp till 10 miljoner tokens och möjliggör lokal körning av enorma modeller på konsumenthårdvara via effektiv MoE-teknik.
Påverkar det EU?
Ja, modellerna släpps med öppen licens och kan laddas ned och användas av utvecklare och företag i EU.
Original source
Entity-watch: Meta AI·36kr.com

The link opens in a new window and leads to the publisher's own site.

Verifierad signal

Källan har spårats automatiskt från utgivaren via Aheadlines signalkedja.

AI-verktyg i artikeln

Topics

#Open Source#AI-benchmarking#Large Language Models (LLMs)#Multimodal AI#Models#Meta AI
[ STAY UP TO DATE ]

Get similar news straight to your inbox

No affiliate linksCancel anytimeGDPR-friendly
[ Frequency ]
[ What do you want to read about? ]

You'll receive updates on 2 topics.

The reader's room

Send in a question or an addition. The newsroom reads everything before it's published and replies when relevant. No AI-generated text – just people.

Sign in to submit a comment or question.

Loading comments…
How this affects you

Read the article through your role

  • Decide whether this affects strategy over 6–12 months or is just noise.
  • Discuss with leadership: do we own the right question or does ownership need to move?
  • Ask: what risk are we taking by NOT acting on this this quarter?

Generated angle — not editorial analysis of "Meta Launches Llama 4: MoE Architecture and Record-Breaking "