IBM released Granite 4.2, a family of dense, decoder-only reasoning LLMs in 3B, 8B, and 30B sizes, all under the Apache 2.0 license. Each model is pre-trained from scratch on roughly 15 trillion tokens, with a five-phase strategy extending the context window to 512K tokens, then fine-tuned and post-trained with multi-stage reinforcement learning. The 8B and 30B models undergo agentic RL in sandboxed environments, and every model includes a thinking/non-thinking switch and native tool calling.
No score is assigned. Sources and their independence are shown in the citation chain below.