Specifications
- Lab
- OpenAI (United States)
- Weights
- Proprietary
- Context window
- 1.1M tokens
- Maximum output
- 128k tokens
- Input / output price
- $2 / $10 per million tokens
- Input modalities
- Image
- Text
- Reasoning
- Yes
- OpenRouter identifier
- openai/gpt-6-sol
At a glance
Updated 5 October 2026
Business Index
98.5%± 1.7 (95% margin of error)
Rank 20 / 27
Average cost per test
$0.018
20th cheapest of 27
Median time
4.1 s
10th fastest of 27
- Average hallucinations
- 0.0%
- Benchmarks taken
- 1 / 1
Scores by business function
The model's mean accuracy across the benchmarks of each business function.
Finance
Not tested
Accounting
98.5%
Human resources
Not tested
Legal
Not tested
Sales
Not tested
Marketing
Not tested
Customer service
Not tested
Procurement & logistics
Not tested
IT
Not tested
Management
Not tested
Results by benchmark
Accounting
- Advertising invoice extraction
98.5%± 1.7
rank 20 of 27
Bars on a 0 to 100% scale. The rank compares the models tested on the same benchmark.
Stability from run to run
A given model alias can change behaviour without notice; re-testing it at every run is how that drift gets caught.
Mean of the model's accuracy across the benchmarks it appears in, run after run. The vertical axis is truncated so that small gaps stay readable.
View the values
| Run | Mean accuracy | Benchmarks |
|---|---|---|
| 2 October 2026 | 100.0% | 1 |
| 4 October 2026 | 97.9% | 1 |
| 5 October 2026 | 98.5% | 1 |