DevPass (LLM Gateway)
Granite 4.2 introduces native reasoning to IBM's open model line, enabling step-by-step planning and self-correction within think tags before answering. The family ships in three sizes, 3B, 8B, and 30B, all released under Apache 2.0, with reasoning switchable across full thinking, non-thinking direct answer, and low-effort brief reasoning modes. This lets a single model skip chain-of-thought on queries where latency matters, paying for reasoning only when needed. Compared to Granite 4.0's fast instruct-only design, the 4.2 family adds reasoning depth aimed at math, multi-step logic, and tool-selection decisions. The model's benchmark gains are most visible in competition math and code tasks, where 8B-level dense reasoning competes effectively. Developers can toggle reasoning effort per request, making Granite 4.2 adaptable for both high-volume low-latency serving and deeper analytical workflows.