OpenAI launches GPT-4o mini, a smaller, cheaper offshoot of GPT-4o, replacing GPT-3.5 Turbo, with support for all multimodal inputs and outputs coming soon
each step more refined Aiming for the ‘perfect training set’ - the ultimate concentrate of all human knowledge and creativity Whether Gemini 1.5 Jeff Harris / @jeffintime : hidden gem: GPT-4o mini supports 16K max_tokens (up from 4K on GPT-4T+GPT-4o) https://openai.com/... [image] Romain Huet / @romainhuet : Such an incredible time to be a developer! GPT-4o mini keeps pushing the cost-intelligence frontier—the cost per token has dropped by 99% since 2022. Model launches at @OpenAI are always special, and this one happens to be on my birthday! Thanks for the gift of GPT-4o mini! 🎉 Gary Marcus / @garymarcus : OpenAI's price cuts are right in line with predictions #2, #3, #4, and #7 below; all seven still look to be on track. As I warned last August, generative AI may turn out to be a dud. Mira Murati / @miramurati : GPT-4o mini makes intelligence far more affordable opening up a wide range of applications. Available on API and rolling out on ChatGPT today Greg Brockman / @gdb : We built gpt-4o mini due to popular demand from developers. We ❤️ developers, and aim to provide them the best tools to convert machine intelligence into positive applications across every domain. Please keep the feedback coming. Ethan Mollick / @emollick : First impressions with GPT-4o-mini (what a name) is that it is impressive for a small model but no replacement for a frontier model. When given complex education prompts it can't follow instructions as well & misses nuance GPT-4o nails I hope it will not be the only free model. [image] @luke_metro : Guys I don't think AI is going to destroy our power grid Romain Huet / @romainhuet : Launching GPT-4o mini: the most intelligent and cost-efficient small model yet! Smarter and cheaper than GPT-3.5 Turbo, it's ideal for function calling, large contexts, real-time interactions—and has vision capabilities. Can't wait to see what you build! https://openai.com/... Andrew Curran / @andrewcurran_ : For the many people asking for the API pricing for GPT-4o Mini it is: 15¢ per M token input 60¢ per M token output 128k context window Andrew Curran / @andrewcurran_ : According to The Verge MMLU benchmarks look like this: GPT-4o - 88.7 GPT-4o Mini - 82 GPT-3.5- 70 https://x.com/... [image] Andrew Curran / @andrewcurran_ : API pricing from TechCrunch. 15 cents/M input 60 cents/M output Context length 128k [image] Andrew Curran / @andrewcurran_ : From Bloomberg: 'OpenAI also said that GPT-4o mini is the company's first AI model to use a new safety tactic it developed called “instruction hierarchy.” [image] @apples_jimmy : 3.5 getting the boot, finally. Matt Shumer / @mattshumer_ : New @OpenAI model! GPT-4o mini drops today. Seems to be a replacement for GPT-3.5-Turbo (finally!) Seems like this model will be very similar to Claude Haiku — fast / cheap, and very good at handling few-shot prompts [image] Rachel Metz / @rachelmetz : New day, new model; OpenAI is rolling out GPT-4o mini, which will replace GPT-3.5 in ChatGPT. OpenAI Releases GPT-4o Mini, a Cheaper Version of Flagship AI Model https://www.bloomberg.com/... Andrew Curran / @andrewcurran_ : Here's our new model. ‘GPT-4o mini’. ‘the most capable and cost-efficient small model available today’ according to OpenAI. Going live today for free and pro. [image] LinkedIn: Amol Shah : We hear all the time from our customers about the need for faster, cheaper models. — So very excited for the launch of GPT-4o mini, second in our “omni” model family. … Jenny O'Leary : Today, we're announcing GPT-4o mini, our most cost-efficient small model. We expect GPT-4o mini will significantly expand the range … Matt Kroll : OpenAI just launched a new model: We're announcing #GPT-4o mini, our most cost-efficient small model. … Mick Vleeshouwer : Today, OpenAI introduced GPT4o-mini, an affordable, fast, and smart model. — 🚀 Already available for testing on #AzureOpenAI (early access playground) … Forums: r/artificial : One-Minute Daily AI News 7/18/2024 Lobsters : GPT-4o mini
The significance is not merely a new model name: OpenAI is moving ChatGPT and API users away from GPT-3.5 Turbo while making a more capable, multimodal successor the default economical option. That makes model selection increasingly a trade-off among price, latency and task fit rather than a simple flagship-versus-legacy choice.
First-order effects
API customers can shift workloads from GPT-3.5 Turbo to GPT-4o mini, whose published pricing is $0.15 per million input tokens and $0.60 per million output tokens, while retaining text and image input support.
ChatGPT’s lower-cost model slot begins moving from GPT-3.5 Turbo to GPT-4o mini; promised multimodal output support expands the model’s intended use cases once available.
Second-order effects
A cheaper model with newer capabilities raises the performance-and-price bar for small-model offerings, pressuring rival providers to compete on useful-task cost, context capacity and modality support rather than benchmark claims alone.
Developers with high-volume summarization, extraction and support workloads gain a clearer incentive to redesign applications around lower inference costs, while suppliers must support migration from legacy model endpoints.
Third-order effects
If lower-cost models repeatedly absorb capabilities once reserved for flagship tiers, AI application economics will increasingly be set by inference efficiency and workload routing rather than access to a single premium model.
The replacement of a legacy default also gives model buyers more leverage: providers will need to show predictable migration paths, pricing and reliability as customers treat models as interchangeable production inputs.
The trend: This is one data point in the shift toward commoditizing baseline multimodal AI capability while competition concentrates on cost per useful task and operational fit.
Introducing GPT-4o mini! It's our most intelligent and affordable small model, available today in the API. GPT-4o mini is significantly smarter and cheaper than GPT-3.5 Turbo. https://openai.com/... [image]
towards intelligence too cheap to meter: https://openai.com/... 15 cents per million input tokens, 60 cents per million output tokens, MMLU of 82%, and fast. most importantly, we think people will really, really like using the new model.
@karpathy A positive, refreshing example of this was Gemma-2's knowledge distillation from the 27B model into the smaller versions (and also MiniLM before that). However, it's also true that MMLU is just multiple-choice answering. It's a good indicator of knowledge, but as we all…
This is not very different from Tesla with self-driving networks. What is the “offline tracker” (presented in AI day)? It is a synthetic data generating process, taking the previous, weaker (or e.g. singleframe, or bounding box only) models, running them over clips in an offline
GPT-4o Mini, announced today, is very impressive for how cheap it is being offered 👀 With a MMLU score of 82% (reported by TechCrunch), it surpasses the quality of other smaller models including Gemini 1.5 Flash (79%) and Claude 3 Haiku (75%). What is particularly exciting is [im…
Make big models, make good data, make models smoler It's like fractional distillation of data instead of oil It's a staircase — each step more refined Aiming for the ‘perfect training set’ - the ultimate concentrate of all human knowledge and creativity Whether Gemini 1.5
Such an incredible time to be a developer! GPT-4o mini keeps pushing the cost-intelligence frontier—the cost per token has dropped by 99% since 2022. Model launches at @OpenAI are always special, and this one happens to be on my birthday! Thanks for the gift of GPT-4o mini! 🎉
First impressions with GPT-4o-mini (what a name) is that it is impressive for a small model but no replacement for a frontier model. When given complex education prompts it can't follow instructions as well & misses nuance GPT-4o nails I hope it will not be the only free model. […
OpenAI's price cuts are right in line with predictions #2, #3, #4, and #7 below; all seven still look to be on track. As I warned last August, generative AI may turn out to be a dud.
LLM model size competition is intensifying… backwards! My bet is that we'll see models that “think” very well and reliably that are very very small. There is most likely a setting even of GPT-2 parameters for which most people will consider GPT-2 “smart”. The reason current …
We built gpt-4o mini due to popular demand from developers. We ❤️ developers, and aim to provide them the best tools to convert machine intelligence into positive applications across every domain. Please keep the feedback coming.
Launching GPT-4o mini: the most intelligent and cost-efficient small model yet! Smarter and cheaper than GPT-3.5 Turbo, it's ideal for function calling, large contexts, real-time interactions—and has vision capabilities. Can't wait to see what you build! https://openai.com/...
From Bloomberg: 'OpenAI also said that GPT-4o mini is the company's first AI model to use a new safety tactic it developed called “instruction hierarchy.” [image]
New @OpenAI model! GPT-4o mini drops today. Seems to be a replacement for GPT-3.5-Turbo (finally!) Seems like this model will be very similar to Claude Haiku — fast / cheap, and very good at handling few-shot prompts [image]
New day, new model; OpenAI is rolling out GPT-4o mini, which will replace GPT-3.5 in ChatGPT. OpenAI Releases GPT-4o Mini, a Cheaper Version of Flagship AI Model https://www.bloomberg.com/...
Here's our new model. ‘GPT-4o mini’. ‘the most capable and cost-efficient small model available today’ according to OpenAI. Going live today for free and pro. [image]