Specifications
- Lab
- Alibaba (China)
- Weights
- Undisclosed
- Context window
- 1M tokens
- Maximum output
- 131k tokens
- Input / output price
- $2 / $6 per million tokens
- Input modalities
- Text
- Image
- Video
- Reasoning
- Yes
- OpenRouter identifier
- qwen/qwen3.8-max-0902
At a glance
Updated 5 October 2026
Business Index
99.3%± 1.2 (95% margin of error)
Rank 15 / 27
Average cost per test
$0.020
21st cheapest of 27
Median time
14.8 s
22nd fastest of 27
- Average hallucinations
- 0.0%
- Benchmarks taken
- 1 / 1
Scores by business function
The model's mean accuracy across the benchmarks of each business function.
Finance
Not tested
Accounting
99.3%
Human resources
Not tested
Legal
Not tested
Sales
Not tested
Marketing
Not tested
Customer service
Not tested
Procurement & logistics
Not tested
IT
Not tested
Management
Not tested
Results by benchmark
Accounting
- Advertising invoice extraction
99.3%± 1.2
rank 15 of 27
Bars on a 0 to 100% scale. The rank compares the models tested on the same benchmark.
Stability from run to run
A given model alias can change behaviour without notice; re-testing it at every run is how that drift gets caught.
Mean of the model's accuracy across the benchmarks it appears in, run after run. The vertical axis is truncated so that small gaps stay readable.
View the values
| Run | Mean accuracy | Benchmarks |
|---|---|---|
| 2 October 2026 | 100.0% | 1 |
| 4 October 2026 | 98.9% | 1 |
| 5 October 2026 | 99.3% | 1 |