Specifications
- Lab
- Alibaba (China)
- Weights
- Proprietary
- Context window
- 1M tokens
- Maximum output
- 131k tokens
- Input / output price
- $4 / $12 per million tokens
- Input modalities
- Text
- Image
- Video
- Reasoning
- Yes
- OpenRouter identifier
- qwen/qwen3.8-max-primeDated version, as of 8 October 2026: qwen/qwen3.8-max-prime-20260923
At a glance
Updated 8 October 2026
Business Index
99.7%± 0.6 (95% margin of error)
Rank 5 / 27
54.8%100.0%
Average cost per test
$0.026
25th cheapest of 27
$0.0002$0.058
Median time
8.6 s
19th fastest of 27
1.5 s29.1 s
- Average hallucinations
- 0.0%
- Benchmarks taken
- 2 / 2
Scores by business function
The model's mean accuracy across the benchmarks of each business function.
Finance
100.0%
Accounting
99.3%
Human resources
Not tested
Legal
Not tested
Sales
Not tested
Marketing
Not tested
Customer service
Not tested
Procurement & logistics
Not tested
IT
Not tested
Management
Not tested
Results by benchmark
Finance
- Financial report analysis
100.0%± 0.0
rank 5 of 27
Accounting
- Advertising invoice extraction
99.3%± 1.2
rank 18 of 27
Bars on a 0 to 100% scale. The rank compares the models tested on the same benchmark.
Stability from run to run
A given model alias can change behaviour without notice; re-testing it at every run is how that drift gets caught.
No benchmark has been published twice with this model yet: there is no curve to draw.