/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

OpenAI launches GPT-4o mini, a smaller, cheaper offshoot of GPT-4o, replacing GPT-3.5 Turbo, with support for all multimodal inputs and outputs coming soon

each step more refined Aiming for the ‘perfect training set’ - the ultimate concentrate of all human knowledge and creativity Whether Gemini 1.5 Jeff Harris / @jeffintime : hidden gem: GPT-4o mini supports 16K max_tokens (up from 4K on GPT-4T+GPT-4o) https://openai.com/... [image] Romain Huet / @romainhuet : Such an incredible time to be a developer! GPT-4o mini keeps pushing the cost-intelligence frontier—the cost per token has dropped by 99% since 2022. Model launches at @OpenAI are always special, and this one happens to be on my birthday! Thanks for the gift of GPT-4o mini! 🎉 Gary Marcus / @garymarcus : OpenAI's price cuts are right in line with predictions #2, #3, #4, and #7 below; all seven still look to be on track. As I warned last August, generative AI may turn out to be a dud. Mira Murati / @miramurati : GPT-4o mini makes intelligence far more affordable opening up a wide range of applications. Available on API and rolling out on ChatGPT today Greg Brockman / @gdb : We built gpt-4o mini due to popular demand from developers. We ❤️ developers, and aim to provide them the best tools to convert machine intelligence into positive applications across every domain. Please keep the feedback coming. Ethan Mollick / @emollick : First impressions with GPT-4o-mini (what a name) is that it is impressive for a small model but no replacement for a frontier model. When given complex education prompts it can't follow instructions as well & misses nuance GPT-4o nails I hope it will not be the only free model. [image] @luke_metro : Guys I don't think AI is going to destroy our power grid Romain Huet / @romainhuet : Launching GPT-4o mini: the most intelligent and cost-efficient small model yet! Smarter and cheaper than GPT-3.5 Turbo, it's ideal for function calling, large contexts, real-time interactions—and has vision capabilities. Can't wait to see what you build! https://openai.com/... Andrew Curran / @andrewcurran_ : For the many people asking for the API pricing for GPT-4o Mini it is: 15¢ per M token input 60¢ per M token output 128k context window Andrew Curran / @andrewcurran_ : According to The Verge MMLU benchmarks look like this: GPT-4o - 88.7 GPT-4o Mini - 82 GPT-3.5- 70 https://x.com/... [image] Andrew Curran / @andrewcurran_ : API pricing from TechCrunch. 15 cents/M input 60 cents/M output Context length 128k [image] Andrew Curran / @andrewcurran_ : From Bloomberg: 'OpenAI also said that GPT-4o mini is the company's first AI model to use a new safety tactic it developed called “instruction hierarchy.” [image] @apples_jimmy : 3.5 getting the boot, finally. Matt Shumer / @mattshumer_ : New @OpenAI model! GPT-4o mini drops today. Seems to be a replacement for GPT-3.5-Turbo (finally!) Seems like this model will be very similar to Claude Haiku — fast / cheap, and very good at handling few-shot prompts [image] Rachel Metz / @rachelmetz : New day, new model; OpenAI is rolling out GPT-4o mini, which will replace GPT-3.5 in ChatGPT. OpenAI Releases GPT-4o Mini, a Cheaper Version of Flagship AI Model https://www.bloomberg.com/... Andrew Curran / @andrewcurran_ : Here's our new model. ‘GPT-4o mini’. ‘the most capable and cost-efficient small model available today’ according to OpenAI. Going live today for free and pro. [image] LinkedIn: Amol Shah : We hear all the time from our customers about the need for faster, cheaper models.  —  So very excited for the launch of GPT-4o mini, second in our “omni” model family. … Jenny O'Leary : Today, we're announcing GPT-4o mini, our most cost-efficient small model.  We expect GPT-4o mini will significantly expand the range … Matt Kroll : OpenAI just launched a new model: We're announcing #GPT-4o mini, our most cost-efficient small model. … Mick Vleeshouwer : Today, OpenAI introduced GPT4o-mini, an affordable, fast, and smart model.  —  🚀 Already available for testing on #AzureOpenAI (early access playground) … Forums: r/artificial : One-Minute Daily AI News 7/18/2024 Lobsters : GPT-4o mini

CNBC Hayden Field

Context & Ripple Effects

GPT-4o mini follows OpenAI’s May introduction of GPT-4o as its faster, natively multimodal flagship, extending that model family down into a lower-cost API tier. It also continues OpenAI’s established pattern of pairing capability updates with lower token prices, seen in the cheaper GPT-4 Turbo API release.

The significance is not merely a new model name: OpenAI is moving ChatGPT and API users away from GPT-3.5 Turbo while making a more capable, multimodal successor the default economical option. That makes model selection increasingly a trade-off among price, latency and task fit rather than a simple flagship-versus-legacy choice.

First-order effects

  • API customers can shift workloads from GPT-3.5 Turbo to GPT-4o mini, whose published pricing is $0.15 per million input tokens and $0.60 per million output tokens, while retaining text and image input support.
  • ChatGPT’s lower-cost model slot begins moving from GPT-3.5 Turbo to GPT-4o mini; promised multimodal output support expands the model’s intended use cases once available.

Second-order effects

  • A cheaper model with newer capabilities raises the performance-and-price bar for small-model offerings, pressuring rival providers to compete on useful-task cost, context capacity and modality support rather than benchmark claims alone.
  • Developers with high-volume summarization, extraction and support workloads gain a clearer incentive to redesign applications around lower inference costs, while suppliers must support migration from legacy model endpoints.

Third-order effects

  • If lower-cost models repeatedly absorb capabilities once reserved for flagship tiers, AI application economics will increasingly be set by inference efficiency and workload routing rather than access to a single premium model.
  • The replacement of a legacy default also gives model buyers more leverage: providers will need to show predictable migration paths, pricing and reliability as customers treat models as interchangeable production inputs.

The trend: This is one data point in the shift toward commoditizing baseline multimodal AI capability while competition concentrates on cost per useful task and operational fit.

Discussion

  • @openaidevs @openaidevs on x
    Introducing GPT-4o mini! It's our most intelligent and affordable small model, available today in the API. GPT-4o mini is significantly smarter and cheaper than GPT-3.5 Turbo. https://openai.com/... [image]
  • @sama Sam Altman on x
    towards intelligence too cheap to meter: https://openai.com/... 15 cents per million input tokens, 60 cents per million output tokens, MMLU of 82%, and fast. most importantly, we think people will really, really like using the new model.
  • @openai @openai on x
    We're continuing to make advanced AI accessible to all with the launch of GPT-4o mini, now available in the API and rolling out in ChatGPT today.
  • @rasbt Sebastian Raschka on x
    @karpathy A positive, refreshing example of this was Gemma-2's knowledge distillation from the 27B model into the smaller versions (and also MiniLM before that). However, it's also true that MMLU is just multiple-choice answering. It's a good indicator of knowledge, but as we all…
  • @karpathy Andrej Karpathy on x
    This is not very different from Tesla with self-driving networks. What is the “offline tracker” (presented in AI day)? It is a synthetic data generating process, taking the previous, weaker (or e.g. singleframe, or bounding box only) models, running them over clips in an offline
  • @tobi Tobi Lutke on x
    16k output tokens is an significant update
  • @luke_metro @luke_metro on x
    Guys I don't think AI is going to destroy our power grid
  • @artificialanlys @artificialanlys on x
    GPT-4o Mini, announced today, is very impressive for how cheap it is being offered 👀 With a MMLU score of 82% (reported by TechCrunch), it surpasses the quality of other smaller models including Gemini 1.5 Flash (79%) and Claude 3 Haiku (75%). What is particularly exciting is [im…
  • @elonmusk Elon Musk on x
    @karpathy Yup, same thing happening with real-world AI at Tesla
  • @bilawalsidhu Bilawal Sidhu on x
    Make big models, make good data, make models smoler It's like fractional distillation of data instead of oil It's a staircase — each step more refined Aiming for the ‘perfect training set’ - the ultimate concentrate of all human knowledge and creativity Whether Gemini 1.5
  • @jeffintime Jeff Harris on x
    hidden gem: GPT-4o mini supports 16K max_tokens (up from 4K on GPT-4T+GPT-4o) https://openai.com/... [image]
  • @romainhuet Romain Huet on x
    Such an incredible time to be a developer! GPT-4o mini keeps pushing the cost-intelligence frontier—the cost per token has dropped by 99% since 2022. Model launches at @OpenAI are always special, and this one happens to be on my birthday! Thanks for the gift of GPT-4o mini! 🎉
  • @emollick Ethan Mollick on x
    First impressions with GPT-4o-mini (what a name) is that it is impressive for a small model but no replacement for a frontier model. When given complex education prompts it can't follow instructions as well & misses nuance GPT-4o nails I hope it will not be the only free model. […
  • @garymarcus Gary Marcus on x
    OpenAI's price cuts are right in line with predictions #2, #3, #4, and #7 below; all seven still look to be on track. As I warned last August, generative AI may turn out to be a dud.
  • @miramurati Mira Murati on x
    GPT-4o mini makes intelligence far more affordable opening up a wide range of applications. Available on API and rolling out on ChatGPT today
  • @karpathy Andrej Karpathy on x
    LLM model size competition is intensifying… backwards!  My bet is that we'll see models that “think” very well and reliably that are very very small.  There is most likely a setting even of GPT-2 parameters for which most people will consider GPT-2 “smart”.  The reason current …
  • @gdb Greg Brockman on x
    We built gpt-4o mini due to popular demand from developers. We ❤️ developers, and aim to provide them the best tools to convert machine intelligence into positive applications across every domain. Please keep the feedback coming.
  • @romainhuet Romain Huet on x
    Launching GPT-4o mini: the most intelligent and cost-efficient small model yet! Smarter and cheaper than GPT-3.5 Turbo, it's ideal for function calling, large contexts, real-time interactions—and has vision capabilities. Can't wait to see what you build! https://openai.com/...
  • @andrewcurran_ Andrew Curran on x
    For the many people asking for the API pricing for GPT-4o Mini it is: 15¢ per M token input 60¢ per M token output 128k context window
  • @andrewcurran_ Andrew Curran on x
    According to The Verge MMLU benchmarks look like this: GPT-4o - 88.7 GPT-4o Mini - 82 GPT-3.5- 70 https://x.com/... [image]
  • @andrewcurran_ Andrew Curran on x
    API pricing from TechCrunch. 15 cents/M input 60 cents/M output Context length 128k [image]
  • @andrewcurran_ Andrew Curran on x
    From Bloomberg: 'OpenAI also said that GPT-4o mini is the company's first AI model to use a new safety tactic it developed called “instruction hierarchy.” [image]
  • @apples_jimmy @apples_jimmy on x
    3.5 getting the boot, finally.
  • @mattshumer_ Matt Shumer on x
    New @OpenAI model! GPT-4o mini drops today. Seems to be a replacement for GPT-3.5-Turbo (finally!) Seems like this model will be very similar to Claude Haiku — fast / cheap, and very good at handling few-shot prompts [image]
  • @rachelmetz Rachel Metz on x
    New day, new model; OpenAI is rolling out GPT-4o mini, which will replace GPT-3.5 in ChatGPT. OpenAI Releases GPT-4o Mini, a Cheaper Version of Flagship AI Model https://www.bloomberg.com/...
  • @andrewcurran_ Andrew Curran on x
    Here's our new model. ‘GPT-4o mini’. ‘the most capable and cost-efficient small model available today’ according to OpenAI. Going live today for free and pro. [image]
  • r/artificial r on reddit
    One-Minute Daily AI News 7/18/2024
  • r/OpenAI r on reddit
    GPT-4o mini: advancing cost-efficient intelligence
  • r/LocalLLaMA r on reddit
    New OPENAI Model
  • r/ChatGPT r on reddit
    OpenAI Slashes the Cost of Using Its AI With a “Mini” Model
  • r/singularity r on reddit
    OpenAI debuts mini version of its most powerful model yet