Meta AIGA
Llama 4 Maverick
Meta’s open flagship by default — permissively licensed and multimodal, but no longer leading the open pack.
Read the full Llama 4 Maverick analysisContext
1.0M
Max output
16K
Input /1M
$0.20
Output /1M
$0.70
Live pricing via OpenRouter
Best for
- Permissively-licensed self-hosting in your own VPC
- Native open-weight multimodal (text+image) work
- Long-context reasoning via the 10M-token Scout sibling
Watch out
A data-center model (8x H100 for Maverick FP8), not a consumer-GPU one, and its benchmarks now trail DeepSeek V4, Qwen3.7-Max and Kimi K2.6. The headline 1417 LMArena Elo was an experimental variant, not the public weights.
For creators. Self-host Scout on a single H100 for whole-corpus (10M context) multimodal drafting where you need data control and zero per-token cost.
Benchmarks
| mmlu pro | 80.5 |
| gpqa diamond | 69.8 |
| livecodebench | 43.4 |
Capabilities
- Native text+image multimodal input via early fusion
- 1M-token context (Maverick); Scout sibling offers 10M
- Llama 4 Community License — free commercial use under 700M MAU
- Self-host: FP8 weights need an 8x H100 80GB node (~600GB); Scout (109B/17B) self-hosts on one H100 at ~55GB Q4
- Run via vLLM / Ollama (Scout) / HF, or hosted APIs (~$0.15/$0.60 per 1M)