Granite 4.2: IBM's Open Reasoning Models With a Five-Phase Pre-Training Curriculum and Multi-Stage Agentic RL
IBM just dropped Granite 4.2—3B, 8B, and 30B reasoning models trained from scratch on 15T tokens, with thinking/non-thinking modes and a staged RL pipeline that teaches agents to code, search, and use terminals.