/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

OpenAI launches o3-mini, its latest reasoning model that the company says is largely on par with o1 and o1-mini in capabilities, but runs faster and costs less

OpenAI on Friday launched a new AI “reasoning” model, o3-mini, the newest in the company's o family of reasoning models.

TechCrunch Kyle Wiggers

Context & Ripple Effects

OpenAI had previewed o3 and o3-mini as models designed to “think” before answering, extending the trade-offs introduced by its earlier o1 line between stronger reasoning and higher cost or latency. The launch turns that previously signaled product direction into an available lower-cost tier.

The accompanying decision to give free ChatGPT users access to the model family marks a shift from a specialist capability toward broader distribution: free users can now try OpenAI reasoning models.

First-order effects

  • OpenAI gains a faster, cheaper reasoning-model option positioned near o1 and o1-mini capability, giving ChatGPT and developer use cases a lower-cost route to this class of output.
  • Users who previously had to weigh o1’s reasoning gains against its cost and performance trade-offs get a more economical choice; free ChatGPT access broadens the immediate audience for the capability.

Second-order effects

  • Competing model providers face greater pressure to pair reasoning performance with lower latency and price, rather than treating advanced reasoning as a premium-only feature.
  • For AI teams, model selection becomes more granular: a lower-cost reasoning tier can shift workloads that were uneconomic under the earlier o1 cost-and-performance trade-offs.

Third-order effects

  • If successive releases keep bringing reasoning capability closer to general-purpose model economics, differentiation will move from merely offering reasoning to delivering the best cost per useful task and reliable deployment controls.
  • A broader user base for reasoning models may accelerate product experimentation, while making inference cost management a more central constraint for providers and enterprise buyers.

The trend: Reasoning models are becoming a tiered, increasingly mass-market product category in which speed and inference economics matter alongside benchmark capability.

Discussion

  • @simonwillison.net Simon Willison on bluesky
    Thinking tokens are charged as output tokens (and count towards the output limit) but are completely invisible in the API - you don't even get a summary of them
  • @simonwillison.net Simon Willison on bluesky
    I think the most notable things about o3-mini may be the price - less than half that of GPT-4o - and the output length limit  —  Its output limit is 100,000 tokens (input is 200,000) - compared to 16,000 for GPT-4o and just 8,000 for both R1 and Claude 3.5
  • @marypcbuk Mary Branscombe on bluesky
    all the AI companies who have been making DeepSeek style smaller models but not releasing them (either to avoid undermining their high end models or because customers actually prefer their high end models) will now be flooding the zone. good luck working out which one best solves…
  • @parkerortolani.com Parker Ortolani on bluesky
    DeepSeek's emergence is *hopefully* going to unleash and even greater era of innovation in new software and hardware.  You think Google, OpenAI, and Anthropic have been moving fast these past few years?  You just wait.  The war for AI supremacy is on.  [embedded post]
  • @sama Sam Altman on x
    o3-mini is out! smart, fast model. available in ChatGPT and API. it can search the web, and it shows its thinking. available to free-tier users! click the “reason” button. with ChatGPT plus, you can select “o3-mini-high”, which thinks harder and gives better answers.
  • @theo @theo on x
    This is such a dirty marketing post and I lowkey love it
  • @jadecole2112 Jade Cole on x
    This is cope. These models are not displaying any significant progress forward, and in some tests perform WORSE than their predecessors. @DHuskytron @daaaaashr @Oliver_H_Vagner @TrueAIHound @GaryMarcus @ChombaBupe
  • @garymarcus Gary Marcus on x
    Wait, are you telling me that Sam ... lied?
  • @jadecole2112 Jade Cole on x
    @daaaaashr ... Some folks are showing their own tests that it's about the same or worse than o1 for some tasks. This after OpenAI said this new drop would be “head and shoulders” above the state of the art models.
  • @polynoamial Noam Brown on x
    We at @OpenAI are proud to release o3-mini, including for the FREE tier. On many evals it outperforms o1. We're shifting the entire cost‑intelligence curve. Model intelligence will continue to go up, and the cost for the same intelligence will continue to go down. [image]
  • @wongmjane Jane Manchun Wong on x
    finalversion_FINAL_v3(2).psd [image]
  • @lexfridman Lex Fridman on x
    OpenAI o3-mini is a good model, but DeepSeek r1 is similar performance, still cheaper, and reveals its reasoning. Better models will come (can't wait for o3pro), but the “DeepSeek moment” is real. I think it will still be remembered 5 years from now as a pivotal event in tech
  • @sama Sam Altman on x
    o3-mini-high is really good and it's cool to hear some people say it's their favorite model ever but it reminds me of when we used to have something internally called “the little big run” a top 2025 goal is to fix our naming problem
  • @ashtom Thomas Dohmke on x
    Today, we bring o3-mini to GitHub Copilot and Models. Small but mighty ftw https://github.blog/...
  • @arankomatsuzaki Aran Komatsuzaki on x
    The leap from o1 to o3 is exponential, completely bypassing o2. If this pattern holds, o3 won't lead to o4—it'll jump straight to o9. [image]
  • @yacinemtb Kache on x
    no chain of thought tokens? [image]
  • @iscienceluvr Tanishq Mathew Abraham, Ph.D. on x
    Sam Altman admits “we will maintain less of a lead than we did in previous years” due to recent developments like DeepSeek. [image]
  • @openai @openai on x
    Try search + reasoning together in ChatGPT. Free users can use OpenAI o3-mini with search by selecting the Search + Reason buttons together. Paid users can select o3-mini from the model picker. [image]
  • @mattshumer_ Matt Shumer on x
    Some initial impressions of o3 mini: - it's clear that the benchmarks don't fully capture how good this model is — it's clearly the best model I've used, for code - the Cursor team has not figured out how to get it to work well in Composer — ChatGPT gives far better results
  • @kyliebytes Kylie Robison on x
    ive been wanting search + reasoning - excited to try it out
  • @simonw Simon Willison on x
    I think the most notable things about o3-mini may be the price - less than half that of GPT-4o - and the output length limit Its output limit is 100,000 tokens (input is 200,000) - compared to 16,000 for GPT-4o and just 8,000 for both R1 and Claude 3.5
  • @zoink Dylan Field on x
    Fun case where o1 pro-mode fails but o3-mini passes: making a “recursive prompt worm” (prompt generates a response that can be used as a prompt to generate a similar — but distinct — response, and so on.) Open challenge for X: reply with a prompt that reliably generates a
  • @pitdesi Sheel Mohnot on x
    Letting users choose between 7 models can't be the right UX for this 90% of users don't know the difference between any of them... at least they could include some descriptors! [image]
  • @firstadopter Tae Kim on x
    DeepSeek - for all the misleading narratives and claims - is a 10 times better name than nearly every U.S. AI model. Simple, direct, and memorable. Why are U.S. companies so terrible at naming products? Stop using consultants and use your gut instinct instead.
  • @emollick Ethan Mollick on x
    Two weeks ago, reasoning models (strongest at hard problems) were only available for subscription Now, if you don't want to pay, you can get reasoners for free from Microsoft Copilot (o1), ChatGPT (o3-mini), or DeepSeek (r1). The best reasoners (o1-pro & o3-mini-high) still cost
  • @miles_brundage Miles Brundage on x
    There continues to be a lot of alpha in reading system cards. So much important stuff going on in the past few months of red teaming/safety mitigations, some concerning (wild capabilities, deception) + encouraging (much better compliance w rules on avg.) https://x.com/...
  • @openai @openai on x
    Try search + reasoning together in ChatGPT. Free users can use OpenAI o3-mini with search by selecting the Search + Reason buttons together. Paid users can select o3-mini from the model picker. [image]
  • @miles_brundage Miles Brundage on x
    Few things are clearer evidence for Moravec's paradox than a small model (o3-mini) just crushing most humans on many things humans think are super hard reasoning problems... ...while lacking a lot of real world knowledge/commonsense + not being able to do any sensorimotor stuff
  • @cursor_ai @cursor_ai on x
    o3-mini is out to all Cursor users! We're launching it for free for the time being, to let people get a feel for the model. The Cursor devs still prefer Sonnet for most tasks, which surprised us.
  • @yuchenj_uw Yuchen Jin on x
    o3-mini might be the best LLM for real-world physics. Prompt: “write a python script of a ball bouncing inside a tesseract” [video]
  • @sama Sam Altman on x
    check out how more details here, especially worth noting the performance of high. also, team really delivered on API pricing! https://openai.com/...
  • @openai @openai on x
    OpenAI o3-mini also works with search to find up-to-date answers with links to relevant web sources. This is an early prototype as we work to integrate search across our reasoning models.
  • @openai @openai on x
    All paid users also have the option of selecting ‘o3-mini-high’ in the model picker for a higher-intelligence version that takes a little longer to generate responses. Pro users will have unlimited access to o3-mini-high.
  • @caromcc_ Caroline McCloskey on x
    https://openai.com/... our first reasoning model that's production-ready out the gate. o3-mini brings function calling, Structured Outputs, developer messages and “reasoning effort.” available now on our api.
  • @nickadobos Nick Dobos on x
    Initial vibe check impression of o3-mini in cursor after playing around for an hour: Not as good as Sonnet 3.5 (new) o1-pro still feels better than both o3-mini definitely hallucinating some swift APIs. That being said, its pretty fast, and I think comparable or better than
  • @emollick Ethan Mollick on x
    Also, obligatory OpenAI-can't-name-a-model comment: o3-mini-high and o3-mini, really?
  • @iscienceluvr Tanishq Mathew Abraham, Ph.D. on x
    idk this does not seem very promising [image]
  • @mckaywrigley Mckay Wrigley on x
    I replaced every ai agent & workflow I have running OpenAI's o1 model with new o3-mini. They *all* still work, and some work *better*. But for 9x cheaper and 4x faster. They are significantly under-hyping this model - it's absolutely unbelievable. o3 & o3 pro will be crazy. [imag…
  • @paulgauthier Paul Gauthier on x
    o3-mini scored similarly to o1 at 10X less cost on the aider polyglot benchmark (both high reasoning). 62% $186 o1 high 60% $18 o3-mini high 54% $9 o3-mini medium https://aider.chat/... [image]
  • @windsurf_ai @windsurf_ai on x
    o3-mini is now available in Windsurf! [video]
  • @iruletheworldmo @iruletheworldmo on x
    i asked. and he is including o1 pro here. so: no model, including o1 pro, can get this right (even with multiple attempts). o3 mini one shots it. intelligence rockets up prices tumbles down.
  • @slow_developer Haider on x
    VERY IMPORTANT - o3-mini (high reasoning) is way better at solving tough math problems than the o1-mini/o1 - o3-mini can solve more than 28% of T3 problems now, imagine what the full o3 could achieve with it? 50%? o4 is already in training, and this benchmark might be fully [imag…
  • @sama Sam Altman on x
    a lot of people prefer this to o1, and it's just the mini model. now we work on the big brother.
  • @brianryhuang Brian Huang on x
    If anyone is curious about Humanity's Last Exam scores: 11.2% on high reasoning, 8.5% on medium reasoning, 5.4% on low reasoning
  • @flowersslop Flowers on x
    o3-mini does not convince me. First impression somewhat disappointing. It is not materially better than r1, and sometimes even worse.
  • @swyx @swyx on x
    ## DeepSeek's impact on o3-mini and o1-mini buried in today's announcement is a 63% (2.7x) price cut for o1 mini - and o3-mini is priced the same. This is much lower than the 25x price cut needed/predicted to match the DeepSeek R1/v3 price curve. In other words, the “USA [image]
  • @dorialexander Alexander Doria on x
    Really insane vibe shift in two weeks.
  • @garymarcus Gary Marcus on x
    𝗛𝗼𝘁 𝘁𝗮𝗸𝗲 𝗼𝗻 𝗼𝟯-𝗺𝗶𝗻𝗶-𝗵𝗶𝗴𝗵 𝗲𝘁𝗰: Another day, another new model, and I am already a few hours later getting bug reports. (Feel free to add your own below.) Folks, they don't call this stuff “GPT-5” because it isn't up to what people have come to expect of
  • @rhydhimma @rhydhimma on x
    Tried o3-mini-high on the same problem. And Nope... It couldn't crack this one. These models are still bad at math problems that require visual solutions. @GaryMarcus
  • @garymarcus Gary Marcus on x
    Another day, another new model, and I am already a few hours later getting bug reports. (Feel free to add your own below.) Folks, they don't call this stuff “GPT-5” because it isn't up to what people have come to expect of that name. It's a better GPT-4 but not the magic
  • @zeroshotnothing Michael Fraser on x
    o3 mini thoughts: price is insane speed is insane latency is insane Just gotta see how much weight it can really pull - can it be as smart or smarter than sonnet with these advantages???
  • @emollick Ethan Mollick on x
    o3-mini-high seems very “smart,” even leaving aside that it is a mini model. Claude and o1 were the only models that could do a sestina one-shot before. [image]
  • @iruletheworldmo @iruletheworldmo on x
    o3 mini is the most capable (publicly available) reasoning model in the world.
  • @maxwinebach Max Weinbach on x
    So o3-mini pre mitigation and safety turning was... the best money at extracting money out of people as a con-artist (when tested against another model as the person) [image]
  • @daniel_mac8 Dan Mac on x
    🔥 as someone that uses LLMs mainly for coding most interested in these massive gains o3-mini with tools has made on SWE-Bench verified 🤯 [image]
  • @scaling01 @scaling01 on x
    GPT-4o is literally better than o3-mini for internal pull requests [image]
  • @danshipper Dan Shipper on x
    We've had o3-mini internally @every for a few days. Here's the vibe check from Cora GM @kieranklaassen: pros: It's like o1 but faster cons: Sonnet still better for html, css, and UX
  • @aaronlopes2 Aaron Lopes on x
    o3-mini-high is absurdly fast at codegen and reasoning. so excited to build with this; phenomenal @OpenAI
  • @ai_for_success AshutoshShrivastava on x
    It's really good from OpenAI to provide API access for o3-mini. Note : o3-mini does not support vision capabilities, But it supports features like function calling, Structured Outputs and developer messages and you can choose between three reasoning effort options, low, [image]
  • @kieranklaassen Kieran Klaassen on x
    o3-mini dropped! While using it over the last week, I learned that @OpenAI is the nerd who does things technically perfectly. @AnthropicAI is the artist with more fun and creative outputs and flair. Both are useful!
  • @jakehalloran1 Jake Halloran on x
    adding o3-mini-high as a separate option from 03-mini is the funniest thing they have ever done. thats just the stupidest thing ever. every company microsoft touches cannot be trusted to name things
  • @shyamalanadkat Shyamal on x
    o3-mini examples thread
  • @_kevinlu Kevin Lu on x
    We released o3-mini, available today to all users in ChatGPT (for free)! o3-mini-low is faster (and often better) than o1-mini, and o3-mini-high is the most capable publicly available reasoning model in the world. https://openai.com/... with @ren_hongyu @shengjia_zhao [image]
  • @andrewwhite01 @andrewwhite01 on x
    Organic chemistry professors remain safe. o3-mini still cannot name organic molecules. It confused cholesterol-3-13C with lanosterol and got the IUPAC name wrong. Different behavior than o1 though - a lot more reasoning cycles of reflecting/updating response. [image]
  • @emollick Ethan Mollick on x
    o3-mini-high does this challenge 1st first time (no other model has made working shaders in many tries, let alone one) “create a visually interesting shader that can run in twigl-dot-app make it like the ocean in a storm...” “make it even more interesting” Scary good model. [imag…
  • @morqon Morgan on x
    six months later: we pass 60% on SWE-bench verified, o3-mini with an internal tool scaffold gets 61% [image]
  • @northstarbrain Alex Northstar on x
    o3-mini-high? What the hell is that? [image]
  • @nicdunz Nic on x
    based on the prices of o1-mini and o3-mini and the fact that plus users got 50 messages per day of o1-mini, chatgpt did the math and decided that it would be best to give plus users 135 to 140 messages per day of o3-mini. we are actually getting 150 messages per day of o3-mini [i…
  • @sherwinwu Sherwin Wu on x
    Best part of launching o3-mini in the API: - available to our broadest user base yet: tiers 3 and up - comes with almost all developer features enabled - cheapest reasoning model yet - smarter than even o1 at times (at reasoning_effort: high)
  • @zan2434 Zain Shah on x
    Most surprising new capability from the o3-mini system card. It looks like o3-mini (in pink - before safety mitigations) randomly got extremely good at conning 4o for money 💸 [image]
  • @scaling01 @scaling01 on x
    o3-mini is insanely cracked at math but that is to be expected math RL is easier than code RL [image]
  • @advaitonline Advait Bopardikar on x
    Silicon Valley is wild, even the AI model are micro-dosing to get more performance “o3-mini-high”
  • @bbacktesting @bbacktesting on x
    o3-mini benchmarked We have a new king! I honestly thought it would do even better because it missed 2 gpt-4o class problems. Intelligence ain't easy. @aidan_mclau @goodside @bindureddy [image]
  • @maxwinebach Max Weinbach on x
    o3-mini-high with web search! Looks like it's making function calls to the search while it's reasoning, and adding the results back into the context window during the reasoning process. I wonder if it's a GPT-4o real time-like model/architecture [image]
  • @theojaffee Theo on x
    A few impressions from the o3-mini system card: [image]
  • @theojaffee Theo on x
    o1 pro mode is still listed as “Best at reasoning”, not o3-mini-high [image]
  • @romainhuet Romain Huet on x
    Welcome, @OpenAI o3-mini! A fantastic model for coding and building agentic apps, complete with function calling, Structured Outputs, and streaming. It outperforms o1 on medium/high reasoning tasks, is much faster, and costs 93% less. Can't wait to see what you build!
  • @maxwinebach Max Weinbach on x
    o3-mini has a 200k token context window and 100k token output matches o1 here [image]
  • @theojaffee Theo on x
    Today will be a much bigger day in AI than anyone is expecting. Millions of normies will see the discontinuity between 4o mini and o3-mini for the first time
  • @kieranklaassen Kieran Klaassen on x
    NEW MODELS! running my isometric Forrest generator benchmark on o3-mini! o3-mini is better than o1 pro and r1, but Sonnet 3.5 wins in visuals still. o3-mini is te technically perfect, but not very artistic [image]
  • @nikunjhanda Nikunj Handa on x
    o3-mini is the most feature complete + developer friendly o-series model we've released to date: function calling, structured outputs, streaming, batch, assistants! It's also: 1. It's 90+% cheaper than o1 2. It's _much_ faster 3. And on things like coding, it's better than o1 I
  • @liamfedus William Fedus on x
    All free ChatGPT users now have reasoning models with o3-mini. The cost-intelligence frontier is shifting fast (o3-mini outperforms even o1 on many STEM evals!)
  • @sebastienbubeck Sebastien Bubeck on x
    o3-mini is a remarkable model. Somehow it has *grokked arxiv* in a way that no other model on the planet has, turning it into a valuable research partner! Below is a deceitfully simple question that confuses *all* other models but where o3-mini gives an extremely useful answer! […
  • @nicdunz Nic on x
    uh.. can i never have to use o3-mini low please? its basically o1-mini and worse than o1... [image]
  • @chrisbarber Chris Barber on x
    OpenAI o3-mini System Card is now available, screenshots below: [image]
  • @thexeophon @thexeophon on x
    huh > We suspect o3-mini's low performance is due to poor instruction following and confusion about specifying tools in the correct format [image]
  • @scaling01 @scaling01 on x
    o3-mini System Card [image]
  • @crisgiardina Cristiano Giardina on x
    o3-mini API pricing: Input: $1.10 / 1M tokens ($0.55 cache hit) Output: $4.40 / 1M tokens deepseek r1 (reasoner) pricing: Input: $0.55 / 1M tokens ($0.14 cache hit) Output: $2.19 / 1M tokens
  • @_aidan_clark_ Aidan Clark on x
    o3-mini first try no edits, took 20 sec (told me how to convert to gif too.....) Get excited :) [image]
  • @burkov Andriy Burkov on x
    “o + 3 + mini + high”, OMG. It's like counting in French: 97 is quatre-vingt-dix-sept, so you must calculate in your head 4 x 20 + 10 + 7 to understand the number.
  • @openaidevs @openaidevs on x
    OpenAI o3-mini is now available in the API for developers on tiers 3-5. It comes with a raft of developer features: ⚙️ Function calling 📝 Developer messages 🗂️ Structured Outputs 🧠 Reasoning effort 🌊 Streaming 📦 Batch API support 🤖 Assistants API support
  • @btibor91 Tibor Blaho on x
    o3-mini and o3-mini-high are here [image]
  • @openai @openai on x
    Before deployment, we carefully assessed the safety of OpenAI o3-mini using the same approach to preparedness, external red-teaming, and safety evaluations as OpenAI o1. Similar to o1, we find that o3-mini significantly surpasses GPT-4o on challenging safety and jailbreak
  • @ivanfioravanti Ivan Fioravanti on x
    📣 o3-mini-high and o3-mini win in the bouncing ball in the square! DeepSeek R1 and o1 Pro lost 🤷‍♂️ Perfect! 3 out of 3! Usual prompt: “write a python script for a bouncing yellow ball within a square, make sure to handle collision detection properly. make the square slowly [vide…
  • @cto_junior @cto_junior on x
    OpenAI employees all month: OMG, can't believe how much better o3-mini, won't even touch o1 slop anymore Reality: [image]
  • @bindureddy Bindu Reddy on x
    o3 mini beats o1 which was already beating r1 We will also have it on https://livebench.ai/ as soon as possible, but I am sure we have a new leader!! o3 will be the model to beat!!! Congrats to OpenAI [image]
  • @btibor91 Tibor Blaho on x
    o3-mini will be available to all users via ChatGPT starting Friday - ChatGPT Plus and Team plans - 150 messages per day - ChatGPT Pro subscribers - unlimited access - ChatGPT Enterprise and ChatGPT Edu customers in a week - o3-mini will also be available via OpenAI's API to
  • @rohanpaul_ai Rohan Paul on x
    🔥 OpenAI just dropped o3-mini. Massive performance numbers for Coding → ChatGPT Plus, Team, and Pro users can access o3-mini immediately, with Enterprise access coming within a week. Key Highlights → Pro users will have unlimited access to o3-mini and Plus & Team users [image]
  • @_akhaliq @_akhaliq on x
    OpenAI o3-mini System Card [image]
  • @deryatr_ Derya Unutmaz on x
    I immediately started testing o3-mini, and it's just great, especially with real-time search! The responses are excellent, better than o1 now! Only limitation is attach photo or file is not yet available, hope that's added soon. I'll report more after a detailed comparison.
  • @oliviergodement Olivier Godement on x
    o3-mini is smart and verrry fast. It's my go-to model and I expect it will push the boundaries of high-value use cases enabled by AI.
  • @localghost Aaron Ng on x
    o3-mini (high) beating o1 at software engineering is a win for everyone. The trend is better, faster, cheaper. Excited to see how o3 pro compares. [image]
  • @daniel_mac8 Dan Mac on x
    o3-mini “high” [image]
  • @andrewcurran_ Andrew Curran on x
    o3-mini launch thread. [image]
  • @vedangvatsa @vedangvatsa on x
    🧵 Hidden Gems in the OpenAI o3-mini System Card paper [image]
  • @daniel_mac8 Dan Mac on x
    🪷 o3-mini-high solves the ‘Buddha’ haiku challenge based on this one test... pretty, pretty good!!! (unedited video, not much time rn) 🎥 👇 [video]
  • @borismpower Boris Power on x
    o3-mini is out - this is both the best and fastest reasoning model!
  • @arankomatsuzaki Aran Komatsuzaki on x
    Not sure making o3-mini a standalone release was the best marketing call—there's little reason to pick it over R-1. It's prob pricier as an API, lacks visible CoT, doesn't outperform, isn't open-source, and only seems faster.
  • @_aidan_clark_ Aidan Clark on x
    o3-mini's intelligence x speed combo is incredible, idk what to say other than just try it and see for yourself. This took 8 seconds, how long would it take you? [video]
  • @altryne Alex Volkov on x
    Allright there we go! O3-mini is finally here 🔥 Fast, cheap reasoning model from @OpenAI that supports: - Function Calling - Structured Outputs - Dev Messages - 3 reasoning efforts (low, medium, high) - 24% faster than o1-mini - 100K max output 👏 Same price as o1 mini! 🔥 [image]
  • @openai @openai on x
    OpenAI o3-mini is now available in ChatGPT and the API. Pro users will have unlimited access to o3-mini and Plus & Team users will have triple the rate limits (vs o1-mini). Free users can try o3-mini in ChatGPT by selecting the Reason button under the message composer.
  • @dorialexander Alexander Doria on x
    I think this is the best take on o3: with all theses focus on benchmarks, US AI labs tend to forget models are also product with features and pricing. Cheap API, transparent traces, locally installable, fine-tunable are all good features where o3 fail to compete
  • @altryne Alex Volkov on x
    Got it, also, now we know what the “think” option in the interface is about, o3-mini is launching for free accounts as well 👏 [image]
  • @marvinvonhagen Marvin von Hagen on x
    ‘o3-mini-high’ how is any normal human being supposed to understand what tf that means it you wanna give users choice, it should be on one dimension max (e.g., a slider for faster—smarter) everything else should be hidden behind “more models” @sama [image]
  • @levie Aaron Levie on x
    o3 is now live. Deepseek is so last week. [image]
  • @deanwball Dean W. Ball on x
    I continue to believe that “superhuman persuasion” is probably not a thing.
  • @dorialexander Alexander Doria on x
    Despite R1 and now mini, they still haven't reopened that repo. Sad. https://x.com/...
  • @theojaffee Theo on x
    o3-mini can one-shot over 90% of OpenAI research engineer interview coding questions [image]
  • @victormustar Victor M on x
    OpenAI is one move away from shocking the entire world... and to be honest, I don't see what they have to lose by doing it now 👀 [image]
  • @iscienceluvr Tanishq Mathew Abraham, Ph.D. on x
    o3-mini System Card “OpenAI o3-mini is the latest model in this series. Similarly to OpenAI o1-mini, it is a faster model that is particularly effective at coding. As can be seen in the capability results below, o3-mini surpasses previous models on science (GPQA Diamond), math [i…
  • @dorialexander Alexander Doria on x
    Yeah a new non-code non-data non-paper. [image]
  • @tszzl Roon on x
    really respect deepseek for making a functional, usable website + mobile app + free hosting so that their model actually gets distribution you see a lot of people train very good open models that aren't used by anybody
  • @kimmonismus @kimmonismus on x
    o3-mini + o3-mini high are being released I have added another benchmark. Today, all plus users with o3-mini high will receive a reasonable model that is about 200 Elo points higher than o1 in Codeforce. This is so you can get a feel for how big and important this release is. [im…
  • @minchoi Min Choi on x
    It's happening! OpenAI will drop o3-mini and o3-mini-high today [image]
  • @dylhunn Dylan Hunn on x
    o3-mini is just so good. exceptional intelligence and blazing speed. I haven't been this excited for a model since GPT-4
  • @slow_developer Haider on x
    o3-mini is set to launch today keep these points in mind before launch: - faster than the o1 model - cheaper than o1-mini and about on par with o1 - stronger performance in coding and math - free tier users also get o3-mini - plus tier will get 100 o3-mini queries per day [image]
  • @kimmonismus @kimmonismus on x
    Today OpenAI will launch 2 reasoning models: o3-mini and o3-mini high! Lets freaking go! [image]
  • @slow_developer Haider on x
    o3-mini confirmed for today there's a similar hint of ‘ho ho ho’ just before the reveal of o3 back in dec so we shouldn't be surprised if they have a plan for something more than just o3-mini
  • r/OpenAI r on reddit
    AMA with OpenAI's Sam Altman, Mark Chen, Kevin Weil, Srinivas Narayanan, Michelle Pokrass, and Hongyu Ren
  • r/OpenAI r on reddit
    OpenAI o3-mini
  • r/OpenAI r on reddit
    o3-mini and o3-mini-high are rolling out shortly in ChatGPT
  • r/OpenAI r on reddit
    o3-mini System Card
  • r/programare r on reddit
    OpenAI o3-mini disponibil (cont > =ChatGptPlus )
  • r/singularity r on reddit
    OpenAI o3-mini |  Pushing the frontier of cost-effective reasoning.
  • @simonw Simon Willison on x
    Has anyone seen anything interesting done with that increased output limit yet? o1 has 100,000 as well, o1-mini is 65,536 and o1-preview was only 32,768 (Reasoning tokens come out of that same budget, but I imagine there are prompts that can reason briefly and then output long)
  • @daniel_mac8 Dan Mac on x
    the fact that o3-mini has a 100,000 token output length, and that output length is so much larger in comparison to previous models lends a lot of credence to Ben's idea that the ‘o’ series reasoning models are best seen as ‘brief’ generators rather than chatbots [image]
  • @nickadobos Nick Dobos on x
    The deepseek effect [image]
  • @openai @openai on x
    Free users can now try OpenAI o3-mini in ChatGPT by selecting the Reason button under the message composer. [image]
  • @nickaturley Nick Turley on x
    Reasoning model in free tier!
  • @tomwarren Tom Warren on x
    OpenAI launches new o3-mini reasoning model with a free ChatGPT version. o3-mini should outperform o1 and provide faster, more accurate answers https://www.theverge.com/...