Quality
#49/670
Overall quality score
See how IBM: Granite 4.1 8B compares with other broadly useful AI models across quality, speed, price, context, and practical capabilities.
0
Overall quality score
114.5 t/s
Output tokens per second
$0.05/1M
Per 1M tokens
$0.10/1M
Per 1M tokens
131.1K
Content read at once
IBM: Granite 4.1 8B ranks #49 of 67 for overall quality and #15 for response speed in the current catalogue.
Its listed input price is $0.05/1M, output price is $0.10/1M, and it can work with up to 131.1K of context at once.
Its strongest areas in the current data are coding, everyday help, complex tasks.
Named evaluation records for technical comparison. Results retain their original benchmark names and sources.
| Benchmark | Area | Granite 4.1 8B | Source |
|---|---|---|---|
| Coding Index | coding | 9.5 | OpenRouter / artificial-analysis |
| GPQA | GPQA Diamond | 43.3 | OpenRouter model benchmarks |
| Humanity's Last Exam | HLE | 3.8 | OpenRouter model benchmarks |
| IFBench | IFBench | 38.6 | OpenRouter model benchmarks |
| τ²-Bench Telecom | τ²-Bench Telecom | 27.8 | OpenRouter model benchmarks |
| AA-LCR | AA-LCR | 12 | OpenRouter model benchmarks |
| CritPt | CritPt | 0 | OpenRouter model benchmarks |
| SciCode | SciCode | 21.8 | OpenRouter model benchmarks |
| Terminal-Bench Hard | Terminal-Bench Hard | 0 | OpenRouter model benchmarks |
| AA-Omniscience Accuracy | AA-Omniscience Accuracy | 12.1 | OpenRouter model benchmarks |
| AA-Omniscience Non-Hallucination Rate | AA-Omniscience Non-Hallucination Rate | 12.7 | OpenRouter model benchmarks |
Quality benchmarks
The darker bar marks IBM: Granite 4.1 8B; the lighter bars provide context from other leading models.
Overall capability · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens · Lower is better