Meta's flagship open-weight model. 400B/17B active MoE architecture, 1M token context.
Each score is anchored to human baseline (100). Source URLs link to the original benchmark, leaderboard, or release note.
| Sub-capability | Quality | Source | Score vs. baseline | Score |
|---|---|---|---|---|
| Common Sense Reasoning | vendor reported | ai.meta.com/llama-4 | 66% |
| Sub-capability | Quality | Source | Score vs. baseline | Score |
|---|---|---|---|---|
| Language Understanding | vendor reported | ai.meta.com/llama-4 | 92% |