AI roundup
IBM Granite 4.2: 3B/8B/30B Specs & Benchmarks
Wed 26 August 2026
Dense decoder-only, Apache 2.0, 128K context, agentic RL on 8B/30B
| Parameters | 3B, 8B, 30B (dense decoder-only, non-MoE) |
| Context Window | 128K native (512K config released) |
| SWE-Bench Verified | 47.67 (8B) / 57.00 (30B) — 3B not tested |
| AIME25 (Math) | 78.33 (3B) / 86.67 (8B) / 89.17 (30B) |
| RULER (128K) | 55.30 (3B) / 71.41 (8B) / 81.38 (30B) |
| Pricing | Not disclosed (weights available Apache 2.0) |
Key takeaways
- Agentic RL split: Only 8B and 30B received the agentic reinforcement-learning block for terminal/web tool use; 3B supports tools but lacks specialized training and SWE-Bench scores.
- Training data: 15 trillion pre-training tokens plus 1 trillion synthetic code tokens via CodeAlchemy pipeline.
- Independent verification: No third-party LMSYS or Artificial Analysis replication available yet; all scores above from IBM NeMo Evaluator SDK.
Source: Ars Technica AI