AI Potluck
Model components / Fine-tuned / chat models

o3 / o4-mini

OpenAI

Previous reasoning tier (2025): full-scale o3 plus cost-efficient o4-mini, both using extended chain-of-thought. Powered Codex agent and ChatGPT's coding mode through 2025. Folded into the GPT-5 line in late 2025; OpenAI consolidated the o-series and GPT branches and is no longer maintaining o-series as a separately branded reasoning family. Current reasoning capability lives in GPT-5.x with thinking enabled and in GPT-5.3-Codex.

OpenAI o3 and o4-mini reasoning models, both released Apr 16 2025 (o3-mini preceded Jan 31 2025). Post-trained reasoning/instruct models, API-only + ChatGPT, no downloadable weights. Succeeded by the GPT-5.x reasoning line but still selectable. The o4-mini AIME figure in the capability record is not on either cited source and is flagged. Verified 2026-08-13 via the OpenAI o3/o4-mini system card and the Wikipedia article.

Openness

1 high confidence
1.0
weights
closed
data
closed
code
closed
license
Proprietary(API+ChatGPT only)

Proprietary OpenAI reasoning models; no weights, data, or code released.

Adoption

5 medium confidence
5.0

o3/o4-mini were rolled out as selectable models inside ChatGPT (Plus/Pro and some free) plus the API; ChatGPT reached ~900M WAU (Feb 2026) rising toward ~1B (May 2026). Reach attributed to the ChatGPT surface these models power, not a standalone o3 count. Now superseded as default by GPT-5.x but still used at scale.

  • https://en.wikipedia.org/wiki/OpenAI_o3 recorded 2026-08-13

    "OpenAI released o3-mini to all ChatGPT users (including free-tier) and some API users"; o3 and o3-pro on the paid tiers; no standalone per-model user count published

Capability

4 high confidence
4.0

Frontier-class reasoning at its Apr 2025 launch, state of the art on Codeforces, SWE-bench and MMMU at the time. It has since been surpassed by GPT-5.x, Opus 4.8 (88.6 SWE-bench) and Gemini 3.5, so on a frontier-relative scale it is a 4 rather than a 5. The o4-mini figure of 99.5 does not appear on the Wikipedia source and rests on OpenAI's system card alone; either way the score is set by the o3 figures.

Verified 2026-08-13