Kilo Gateway
> "We also introduced OpenAI o3-pro in the API—a version of o3 that uses ... This announcement adds o3-pro, which pairs with o3 in the same way the o4 ...
Model details
o3-pro belongs to OpenAI's o-series reasoning family and is presented as a higher-compute sibling of the standard o3 model. Rather than introducing a new architecture, it extends the same chain-of-thought approach that the o-series is built on, allocating substantially more inference compute so the model can explore problem-solving paths, check its own logic, and backtrack before producing an answer. This "think longer" design is what distinguishes o3-pro from o3 and makes it especially well suited to hard, multi-step problems where extended deliberation matters more than raw speed.
In practical terms, the model is aimed at workloads where reasoning depth outweighs latency: competition-level mathematics, scientific analysis, complex multi-step code generation, and long-form research assistance. Third-party reporting suggests that on these difficult tasks o3-pro outperforms standard o3, while for routine queries the difference is small and the slower response time is harder to justify. As a result, o3-pro is best framed as a specialist tool for users who need maximum reasoning quality on genuinely hard problems rather than a general-purpose conversational model.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Kilo Gateway
> "We also introduced OpenAI o3-pro in the API—a version of o3 that uses ... This announcement adds o3-pro, which pairs with o3 in the same way the o4 ...
This exact model name is also listed by 6 other providers.