Specifications
- Lab
- Alibaba (China)
- Weights
- Undisclosed
- Context window
- 1M tokens
- Maximum output
- 131k tokens
- Input / output price
- $4 / $12 per million tokens
- Input modalities
- Text
- Image
- Video
- Reasoning
- Yes
- OpenRouter identifier
- qwen/qwen3.8-max-prime
At a glance
Updated 5 October 2026
Business Index
99.3%± 1.2 (95% margin of error)
Rank 18 / 27
Average cost per test
$0.037
25th cheapest of 27
Median time
11.3 s
19th fastest of 27
- Average hallucinations
- 0.0%
- Benchmarks taken
- 1 / 1
Scores by business function
The model's mean accuracy across the benchmarks of each business function.
Finance
Not tested
Accounting
99.3%
Human resources
Not tested
Legal
Not tested
Sales
Not tested
Marketing
Not tested
Customer service
Not tested
Procurement & logistics
Not tested
IT
Not tested
Management
Not tested
Results by benchmark
Accounting
- Advertising invoice extraction
99.3%± 1.2
rank 18 of 27
Bars on a 0 to 100% scale. The rank compares the models tested on the same benchmark.
Stability from run to run
A given model alias can change behaviour without notice; re-testing it at every run is how that drift gets caught.
Mean of the model's accuracy across the benchmarks it appears in, run after run. The vertical axis is truncated so that small gaps stay readable.
View the values
| Run | Mean accuracy | Benchmarks |
|---|---|---|
| 2 October 2026 | 100.0% | 1 |
| 4 October 2026 | 98.9% | 1 |
| 5 October 2026 | 99.3% | 1 |