Llama 4 Maverick
Meta · The open-weights standard-bearer.
Llama 4 anchors the open ecosystem: run it anywhere, fine-tune it freely, pay only for compute — with a 1M-context MoE design.
Context window
1.0M
tokens
Input · $/1M
$0.188
per million tokens
Output · $/1M
$0.652
per million tokens
Vs. cost floor
2×
output cost vs DeepSeek V4.1 Flash
Capability profile
Editorial ratings across six axes (0–100), compiled from publisher reports and public evals — indicative, not measured by TrendNexus.
Benchmark highs
solid open-weights showing
capable coding
multimodal-enabled
most-deployed open family
Where it sits in the field
Output cost per million tokens across all 15 tracked flagships — Llama 4 Maverick highlighted.
What it's suited to build
Product classes where this model's profile wins — mockups are illustrative, not model output.
Self-hosted / edge
Sovereign/self-hosted AI
Regulated data that can never leave your infrastructure.
Product AI
Fine-tuned domain models
Own a specialist model instead of renting a generalist.
Agents
Cost-controlled fleets
Fixed-cost inference for always-on agents.
Benchmark figures and capability ratings are indicative — compiled from publisher reports and public leaderboards, refreshed editorially, and not independently measured by TrendNexus. Verify against your own evals before procurement decisions. Pricing updates with the daily build.