Specifications
- Lab
- Meta (United States)
- Weights
- Open weights
- Context window
- 1M tokens
- Maximum output
- 16k tokens
- Input / output price
- $0.188 / $0.653 per million tokens
- Input modalities
- Text
- Image
- Reasoning
- No
- OpenRouter identifier
- meta-llama/llama-4-maverick
At a glance
Updated 5 October 2026
Business Index
98.3%± 1.8 (95% margin of error)
Rank 22 / 27
Average cost per test
$0.0017
9th cheapest of 27
Median time
5.6 s
13th fastest of 27
- Average hallucinations
- 1.0%
- Benchmarks taken
- 1 / 1
Scores by business function
The model's mean accuracy across the benchmarks of each business function.
Finance
Not tested
Accounting
98.3%
Human resources
Not tested
Legal
Not tested
Sales
Not tested
Marketing
Not tested
Customer service
Not tested
Procurement & logistics
Not tested
IT
Not tested
Management
Not tested
Results by benchmark
Accounting
- Advertising invoice extraction
98.3%± 1.8
rank 22 of 27
Bars on a 0 to 100% scale. The rank compares the models tested on the same benchmark.
Stability from run to run
A given model alias can change behaviour without notice; re-testing it at every run is how that drift gets caught.
Mean of the model's accuracy across the benchmarks it appears in, run after run. The vertical axis is truncated so that small gaps stay readable.
View the values
| Run | Mean accuracy | Benchmarks |
|---|---|---|
| 2 October 2026 | 98.7% | 1 |
| 4 October 2026 | 97.9% | 1 |
| 5 October 2026 | 98.3% | 1 |