OpenAI introduces GPT-4o mini, its most capable and cost-efficient small model, surpassing GPT-3.5 Turbo on academic benchmarks.
OpenAI introduced GPT-4o mini on July 18, 2024 as its most cost-efficient small model[1][2]
Available in the Assistants, Chat Completions and Batch APIs; in ChatGPT Free, Plus and Team users get it that day in place of GPT-3.5, Enterprise the following week[2]
First model in the API to apply the instruction hierarchy method against jailbreaks and prompt injections[2]
OpenAI's release notes list it as the July 18, 2024 entry, later replaced in ChatGPT by GPT-4.1 mini (May 14, 2025)[1]
Notable features
A small model that surpasses GPT-3.5 Turbo and other small models on textual and multimodal academic benchmarks, supports the same languages as GPT-4o and shares its improved tokenizer, with stronger function calling and long-context performance than GPT-3.5 Turbo[3]
It scores 82% on MMLU (Gemini Flash 77.9%, Claude Haiku 73.8%) and 87.2% on HumanEval, and outperforms GPT-4 on chat preferences in the LMSYS leaderboard[2]
It scores 87.0% on MGSM math reasoning (Gemini Flash 75.5%, Claude Haiku 71.7%) and 59.4% on the MMMU multimodal reasoning eval (56.1% and 50.2%)[3]
Priced at 15 cents per million input tokens and 60 cents per million output tokens, more than 60% cheaper than GPT-3.5 Turbo[2]
Context window of 128K tokens, up to 16K output tokens per request, knowledge to October 2023; text and vision in the API with other modalities to come[2]
The first entry is always the pin's source. Overall confidence is a weighted average of how firmly each reference supports the start and end times used above; a reference counts half as much for every 180 days older than the newest.
web.archive.org· Published Jul 18, 2024· 2% of score
Archived copy of OpenAI's[1][2] GPT-4o mini announcement, which says "On MGSM, measuring math reasoning, GPT-4o mini scored 87.0%" and "scoring 59.4%" on MMMU, and that it shares GPT-4o's improved tokenizer.
Suggest a correction
Something missing or wrong? Say it in your own words: a link that backs this pin up, a different start or end date and why, or a fact it lacks or gets wrong. The AI checks it against this pin's sources, searches for better ones, and adds any page that backs you up. The pin's own sources still count most.
CONFIRMED89% CONFIDENCE320 days agoDated from the release notes leave this entry undated; the Codex changelog dates GPT-5-Codex-Mini's introduction Nov 7, 2025[1]
OpenAI adds GPT-5-Codex-Mini, a smaller, cheaper GPT-5-Codex with about 4x more usage, to the Codex CLI and IDE extension.