Specifications
- Lab
- Google (United States)
- Weights
- Open weights
- Context window
- 262k tokens
- Maximum output
- 16k tokens
- Input / output price
- $0.09 / $0.34 per million tokens
- Input modalities
- Image
- Text
- Video
- Reasoning
- Yes
- OpenRouter identifier
- google/gemma-4-31b-itDated version, as of 8 October 2026: google/gemma-4-31b-it-20260402
At a glance
Updated 8 October 2026
Business Index
71.6%± 6.1 (95% margin of error)
Rank 24 / 27
54.8%100.0%
Average cost per test
$0.0002
1st cheapest of 27
$0.0002$0.058
Median time
2.8 s
7th fastest of 27
1.5 s29.1 s
- Average hallucinations
- 0.0%
- Benchmarks taken
- 2 / 2
Scores by business function
The model's mean accuracy across the benchmarks of each business function.
Finance
43.8%
Accounting
99.3%
Human resources
Not tested
Legal
Not tested
Sales
Not tested
Marketing
Not tested
Customer service
Not tested
Procurement & logistics
Not tested
IT
Not tested
Management
Not tested
Results by benchmark
Finance
- Financial report analysis
43.8%± 12.2
rank 24 of 27
Accounting
- Advertising invoice extraction
99.3%± 1.2
rank 7 of 27
Bars on a 0 to 100% scale. The rank compares the models tested on the same benchmark.
Stability from run to run
A given model alias can change behaviour without notice; re-testing it at every run is how that drift gets caught.
No benchmark has been published twice with this model yet: there is no curve to draw.