Granite 4.2 LLMs: Architecture, Training and Multi-Stage RL
Granite 4.2 is a family of dense, decoder-only reasoning LLMs in three sizes (3B, 8B, 30B), pre-trained from scratch on ~15T tokens with a five-phase strategy and a 512K context extension. The 8B and 30B models undergo agentic RL for tool use; all models are Apache 2.0.