Google unveils Gemini 3, its “most intelligent” and “factually accurate” model yet, with improvements across coding and reasoning, and offering less “flattery”
The flagship Gemini 3 Pro model is coming to the Gemini app and Search, with improvements across coding, reasoning, and less ‘flattery.’
The Verge Emma Roth
Related Coverage
- A new era of intelligence with Gemini 3 The Keyword
- Google Search with Gemini 3: Our most intelligent search yet The Keyword · Elizabeth Hamon Reid
- Google's Gemini 3 Is Here: A Special Early Look New York Times
- Google unveils Gemini's next generation, aiming to turn its search engine into a ‘thought partner’ Associated Press · Michael Liedtke
- Google releases Gemini 3 - it already powers AI Mode Search Engine Land · Barry Schwartz
- Google just rolled out Gemini 3 to Search - here's what it can do and how to try it ZDNET · Webb Wright
- Gemini 3 Pro arrives today with fewer lies, more agentic coding Ars Technica · Ryan Whitwam
- Google Seeks to Shake Up Chatbot Race With New Gemini Version Wall Street Journal · Katherine Blunt
- Google announces Gemini 3 as battle with OpenAI intensifies CNBC · Jennifer Elias
- Google launches Gemini 3, embeds AI model into search immediately Reuters · Kenrick Cai
- Google's new Gemini 3 model arrives in AI Mode and the Gemini app Engadget · Igor Bonifacic
- Google just added Gemini 3 to Search - here's who can access it ZDNET · Webb Wright
- Google's Gemini 3 is live after months of hype. Here's what it can do. Business Insider · Hugh Langley
- Google rolls out Gemini 3 Pro to power search and app Axios · Ina Fried
- Google's Gemini 3 AI model is here, available now to Canadians MobileSyrup · Dean Daley
- Google Opens Data Center in Winschoten to Help Power the AI Economy in The Netherlands Google Cloud
- Google unveils Gemini 3 with a commitment to localize data in India Livemint · Shouvik Das
- Google's Gemini Agent can orchestrate complex tasks on your behalf in the Gemini app Android Central · Brady Snyder
- Google releases its heavily hyped Gemini 3 AI in a sweeping rollout—even Search gets it on day one Fortune · Sharon Goldman
- Google launches Gemini 3, Google Antigravity, generative UI features Constellation Research · Larry Dignan
- Three Years from GPT-3 to Gemini 3 One Useful Thing · Ethan Mollick
- Google launches more powerful Gemini 3 model for reasoning and vibe coding The Indian Express · Anuj Bhatia
- Gemini 3 Rumors Are CONFIRMED, It's VERY GOOD Matt Wolfe on YouTube · Matt Wolfe
- Introducing Gemini 3 — Since the Gemini era began, we've been pushing the frontier to build … The Keyword · Sundar Pichai
- I used Google's new Gemini 3 AI to make Android apps and fine-tune my workouts Android Authority · Mishaal Rahman
- Today we're taking another big step on our journey of pushing the frontiers of AI and releasing Gemini 3 - our most intelligent model that helps you bring any idea to life. … Lily Lin
- Google Brings Gemini 3 AI Model to Search and AI Mode Hacker News
- Start building with Gemini 3 The Keyword · Logan Kilpatrick
- Gemini app rolling out Gemini 3 Pro as ‘Gemini Agent’ comes to AI Ultra 9to5Google · Abner Li
- Google's new Gemini 3 vibe-codes its responses and comes with its own agent MIT Technology Review · Caiwei Chen
- Google Antigravity introduces agent-first architecture for asynchronous, verifiable coding workflows VentureBeat · Emilia David
- Antigravity Is Google's New Agentic Development Platform The New Stack · Frederic Lardinois
- Gemini 3 is here: Google's most advanced model promises better reasoning, coding, and more Android Authority · Aamir Siddiqui
- Google AntiGravity: The New Best AI Editor? aiwithbrandon on YouTube · Aiwithbrandon
- Google just dropped their Cursor killer (FREE Gemini 3 Pro???) Theo - t3․gg on YouTube · Theo
- Gemini 3 for developers: New reasoning, agentic capabilities Hacker News
- Google's Gemini 3 AI model makes its long-awaited debut, crushing rivals on top benchmarks SiliconANGLE · Mike Wheatley
- Google unveils Gemini 3 claiming the lead in math, science, multimodal and agentic AI benchmarks VentureBeat · Carl Franzen
- Gemini 3 is here — Google's most powerful AI model yet is crushing benchmarks, improving search and outperforming ChatGPT Tom's Guide · Amanda Caswell
- Google Launches Gemini 3 Pro and Antigravity IDE, Claiming Reasoning Lead Over GPT-5.1 WinBuzzer · Markus Kasanmascheff
- Google Says New Gemini 3 AI Model Will Better Understand Your Requests CNET · Imad Khan
- Gemini 3 Arrives; Open-Source Coding Agent Raises $19 Million The Information · Erin Woo
- Gemini 3 is here - 3 things to know about the major AI update TechRadar · Eric Hal Schwartz
- Gemini 3 is Google's most intelligent AI model and it's available now — here's everything you need to know Android Central · Brady Snyder
- Google AI Mode Turns Search into Travel Planner: What It Means for Fair Competition MediaNama · Aakriti Bansal
- Google seeks AI lead with improved Gemini BSS
- Gemini 3 gets the power to shape Search results for maximum impact Android Authority · Stephen Schenck
- Daily Search Forum Recap: November 18, 2025 Search Engine Roundtable · Barry Schwartz
- Gemini 3 brings upgraded smarts and new capabilities to the Gemini app The Keyword · Josh Woodward
- Thrilled to bring to you Gemini Agent - leveraging Gemini 3.0's capabilities - a project I have been leading since its inception. … Vijay Narayanan
- Google Launches New Gemini AI Model With Interactive Answers Bloomberg
- Gemini 3 Is Here—and Google Says It Will Make Search Smarter Wired · Will Knight
- Google Launches Gemini 3, Claims Benchmark Lead Over GPT-5.1 and Claude Sonnet 4.5 Analytics India Magazine · Siddharth Jindal
Discussion
-
@thoughtlesslabs
@thoughtlesslabs
on x
well gemini 3 is pretty incredible. it still does the thing where if it puts out exactly what you wanted first try - it feels magical. but if you have to make changes, things start to fall apart and you end up in that prompt hell where you start losing things you liked
-
@venturetwins
Justine Moore
on x
Thrilled to find that Gemini 3 has retained one of my favorite features: having an emotional outburst when it can't fix your code 🙃 [image]
-
@minchoi
Min Choi
on x
Ok Gemini 3 is insane. People can't stop just build things 10 wild examples: 1. Theme Park game from 90s remade in hours [video]
-
@sqs
Quinn Slack
on x
I've been using Gemini 3 Pro in Amp. We just shipped it to everyone as the default in ‘smart’ mode. It's much better at finding and using feedback loops, in particular. There are rough edges and more models on the horizon, but it's the best right now and /really/ shines in Amp.
-
@theomediaai
@theomediaai
on x
Google just dropped Gemini 3, and they aren't being subtle. 🤯 They're calling it their “most intelligent model” yet, built on a foundation of state-of-the-art reasoning. It's not just topping benchmarks; it officially surpasses 2.5 Pro at coding, mastering complex zero-shot [vide…
-
@fchollet
François Chollet
on x
Gemini 3 scores 31.1% on ARC-AGI-2. Impressive progress.
-
@thorstenball
Thorsten Ball
on x
Amp's new default model: Gemini 3 Pro https://ampcode.com/... [image]
-
@artemr
Artem Russakovskii
on x
Some of these Gemini 3 performance gains are wild. Unfortunately, it seems that for coding, while much better than Gemini 2.5 Pro, it is in line with existing competition, which actually means it's behind because the competition has surely already started cooking the next thing. …
-
@cgtwts
@cgtwts
on x
Gemini 3.0 is an absolute beast and it can literally do anything. it's over for y'all >chatgpt >claude >grok [image]
-
@antirez
@antirez
on x
Gemini 3 finds the Redis Streams bug I was talking about a few days ago like (and better) GPT 5.1 could. It's a complex bug and yet it creates a clear causal model of what is happening in the crash report.
-
@arcprize
@arcprize
on x
Gemini 3 models from @Google @GoogleDeepMind have made a significant 2X SOTA jump on ARC-AGI-2 (Semi-Private Eval) Gemini 3 Pro: 31.11%, $0.81/task Gemini 3 Deep Think (Preview): 45.14%, $77.16/task [image]
-
@mattshumer_
Matt Shumer
on x
The intelligence of Gemini 3 Deep Think looks to be off-the-charts [image]
-
@emollick
Ethan Mollick
on x
I had access to Gemini 3. It is a very good, very fast model. It also demonstrates the change from chatbot to agent. https://www.oneusefulthing.org/ ...
-
@jacalulu
Jaclyn Konzelmann
on x
Gemini 3 is here... and the vibe coding is unreal. I used it to build my entire new site → https://jaclynkonzelmann.com/ This is the kind of shift you feel instantly as a builder 🤯
-
@rajanpatel
Rajan Patel
on x
Big day for the team: we just introduced Gemini 3 - our most intelligent model - and brought it straight to Google Search on day one, starting with AI Mode. With state-of-the-art reasoning, unparalleled multimodal understanding and powerful coding capabilities, Gemini 3 is built …
-
@amandaksilver
Amanda Silver
on x
Gemini 3 Pro already available in GitHub Copilot! Developer choice so you can do your best work.
-
@clankert800
Abdullah
on x
Gemini 3.0 Pro is insane🔥 > “Make a 3D Rubik's cube simulation, include the buttons “shuffle”, and “solve”. both play a nice smooth animation.” [video]
-
@rihardjarc
Rihard Jarc
on x
If these leaked benchmarks for Gemini 3 Pro are correct (could not verify yet) $GOOGL about to drop a model that dwarfs other models and claims the num 1 spot while being both trained and inferenced on TPUs. $GOOGL about to disrupt the whole AI supply chain with its full stack. […
-
@iterintellectus
Vittorio
on x
happy Gemini 3 pro day to all who celebrate [image]
-
@yitayml
Yi Tay
on x
Gemini 3! This is our most intelligent model that brings any idea to life. 😻 This is the best model in the world, by a crazy wide margin! Aside from a huge increase across the absolutely everything, look at its coding capabilities and quality of aesthetics and fidelity. [video]
-
@deedydas
Deedy
on x
🚨The Gemini 3 launched leaked hours before the official launch!! The model is INSANELY good, but let's dig deep and see what it means for the AI future. 1. It's not going to be super expensive! Google trained this from scratch on TPUs as an MoE with 1M input and 64k token [image]
-
@quocleix
Quoc Le
on x
Gemini 3 is finally out. 🚀 The numbers on the hardest benchmarks are wild. Seeing MathArena Apex go from <2% to 23.4% and ARC-AGI-2 hit 31% feels like a real turning point for reasoning. We're starting to crack problems that used to look impossible. Huge congrats to the team. [im…
-
@googledeepmind
@googledeepmind
on x
🔵 Our best model for vibe and agentic coding Gemini 3 can make beautiful and dynamic apps from a single prompt. It can follow instructions exceptionally well, so it knows exactly what you're asking for - and how to make it. We also improved agentic code performance to support [vi…
-
@skirano
Pietro Schirano
on x
It's also amazing at games. It recreated the old iOS game called Ridiculous Fishing from just a text prompt, including sound effects and music. [video]
-
@ns123abc
Nik
on x
Gemini 3 has destroyed ChatGPT5 [image]
-
@googledeepmind
@googledeepmind
on x
Our first release is Gemini 3 Pro, which is rolling out globally starting today. It significantly outperforms 2.5 Pro across the board: 🥇 Tops LMArena and WebDev @arena leaderboards 🧠 PhD-level reasoning on Humanity's Last Exam 📋 Leads long-horizon planning on Vending-Bench 2 [im…
-
@skirano
Pietro Schirano
on x
It also did something I've seen other LLMs struggle to do before, it built a fully functional Game Boy emulator, and yes, it even drew the Game Boy as an SVG. [image]
-
@googledeepmind
@googledeepmind
on x
🔵 State-of-the-art reasoning It understands and answers prompts with more depth and nuance than ever before. Get clear, direct responses without the clichés and filler. As our most factual model, it's more reliable for complex questions in science and math. [video]
-
@sebkrier
Séb Krier
on x
Gemini 3 is ridiculously good. Two-shot working simulation of a nuclear power plant. Imagine walking through a photorealistic version of this in the next version of Genie! 🧬👽☢️ [video]
-
@googledeepmind
@googledeepmind
on x
This is Gemini 3: our most intelligent model that helps you learn, build and plan anything. It comes with state-of-the-art reasoning capabilities, world-leading multimodal understanding, and enables new agentic coding experiences. 🧵 [video]
-
@googledeepmind
@googledeepmind
on x
🔵 World-leading multimodal understanding It can take in any kind of input you give it - from text to video to code - and responds with whatever best suits your needs. Ask Gemini to break down concepts from a long video or quickly turn a research paper into an interactive guide. […
-
@joshwoodward
Josh Woodward
on x
🚀 Gemini 3 has landed in the @GeminiApp! It's great on the leaderboards and feels great to use. Today's launch also comes with 6 more features 🧵 [video]
-
@elonmusk
Elon Musk
on x
@demishassabis @arena Nice work
-
@demishassabis
Demis Hassabis
on x
Looking forward to seeing what you all build with Gemini 3! Starting today it's rolling out: - In the @GeminiApp for anyone to try - For developers in @GoogleAIStudio, our new agentic development platform Antigravity, and Gemini CLI - For Google AI Pro & Ultra subscribers to use
-
@demishassabis
Demis Hassabis
on x
For example I've been doing a bunch of late night vibe coding with Gemini 3 in @GoogleAIStudio, and it's so much fun! I recreated a testbed of my game Theme Park 🎢 that I programmed in the 90s in a matter of hours, down to letting players adjust the amount of salt on the chips! […
-
@deliprao
@deliprao
on x
I vibe-coded this yesterday with Gemini-3.0-Pro with an extremely low-effort prompt (~200 words, with no punctuation) spread over two turns. The second turn was only because I forgot to add something. Congratulations to @GoogleDeepMind on this incredible model release! To top it …
-
@yuchenj_uw
Yuchen Jin
on x
Holy moly! Gemini 3 Deep Think even outperforms Gemini 3 Pro. - 41% on Humanity's Last Exam - 45.1% on ARC-AGI-2 Google has reclaimed its AI dominance after Transformer. This might be a turning point of an era. Can OpenAI catch up? https://t.co/...
-
@danshipper
Dan Shipper
on x
BREAKING: Gemini 3 Pro is out!! It's state of the art (or close) on coding, reasoning, computer use and more. It's also extremely fast. We've been testing it internally @every for a few hours. Here's what we've noticed so far: - Coding. It rips in @FactoryAI's Droid. So fast [ima…
-
@sriramk
Sriram Krishnan
on x
congratulations to @demishassabis and the team at Google for the launch of Gemini 3.
-
@patrickmoorhead
Patrick Moorhead
on x
Interested in running our company prompts against Gemini 3. We are a Workspace shop and hopefully they made improvements to that interaction. So far, doing things with calendar, in particular, has been weak. I get better outcomes on OpenAI and Perplexity related to GMail and
-
@demishassabis
Demis Hassabis
on x
We've been intensely cooking Gemini 3 for a while now, and we're so excited and proud to share the results with you all. Of course it tops the leaderboards, including @arena, HLE, GPQA etc, but beyond the benchmarks it's been by far my favourite model to use for its style and [im…
-
@officiallogank
Logan Kilpatrick
on x
One of the early Gemini 3 tests I did was take the bouncing ball example and try to make it 10x harder, Gemini 3 Pro crush it in 1 shot... (not best of N, literally first prompt made this) [video]
-
@mattturck
Matt Turck
on x
Gemini 3! Google is back and it's so over for OpenAI and Anthropic, until next week/month of course when it will be so over for Google for at least a few weeks
-
@film_girl
Christina Warren
on x
Huge congrats to the whole GDM team on Gemini 3! So proud of you all!
-
@scaling01
@scaling01
on x
Google is where this journey started with Transformers and where it will end Google has the mandate [image]
-
@lmthang
Thang Luong
on x
#Gemini3 is finally out! Congrats to everyone on this amazing launch! Also very excited to see how #DeepThink can power Gemini3 to further the state-of-the-art performances across reasoning, deep knowledge, and multimodality: 41% HLE, 93.8% GPQA, & 45.1% on ARC-AGI-2 (big jump)! …
-
@rowancheung
Rowan Cheung
on x
Google just released Gemini 3, its most intelligent AI model yet. I caught up with Demis Hassabis, CEO of Google DeepMind, to ask about: -Gemini 3 and Google's AI strategy -Google's new Antigravity tool -Medical-grade AI Here's what he said: https://t.co/...
-
@scaling01
@scaling01
on x
Gemini 3.0 Pro made GPT-5.1 obsolete overnight, it's cheaper and much stronger Gemini 3.0 Pro - 31.1% at $0.811/task GPT-5.1 - 17.6% at $1.17/task [image]
-
@_philschmid
Philipp Schmid
on x
Meet Gemini 3 Pro! Our best coding and agentic Gemini model with state-of-the-art multimodal reasoning! 🚀 Here is all you need to know: - 1M token context window with 64k output, knowledge cut of Jan 2025. - $2 / $12 (<200k tokens); $4 / $18 (>200k tokens). - 30% tool [image]
-
@figma
@figma
on x
Starting today, you will be able to select Gemini 3 Pro as an experimental model in Figma Make It's early, but we're already seeing impressive results: → Generate 0→1 results with a broad variety of styles, layouts, and interactions → Accurately translate Figma designs into https…
-
@kimmonismus
@kimmonismus
on x
Here we go: Gemini 3.0 is live! Thankfully, I was able to be an early tester. It's the best model I've ever tested - nothing comes even close to it. OpenAI will have a very hard time to keep up with Google, because Gemini 3.0 will now be the benchmark for all future models. I [vi…
-
@rmstein
Robby Stein
on x
(1/5) Gemini 3, our most intelligent model, is landing in Google Search today - starting with AI Mode. Excited that this is the first time we're shipping a new Gemini model in Search on day one! 🚀 In Search, Gemini 3 with generative layouts will make it easy to get a rich [video]
-
@patrickc
Patrick Collison
on x
I asked Gemini 3 to make an interactive web page summarizing 10 breakthroughs in genetics over the past 15 years. Result: https://gemini.google.com/... Pretty cool IMO. [image]
-
@omarsar0
Elvis
on x
Gemini 3 is here! The most exciting part of this model is the long horizon planning capabilities. This is going to unlock complex agentic applications, the likes of which we have never seen. Vibe coding insane applications in one shot, proactive agents that expand our [image]
-
@patloeber
Patrick Loeber
on x
Say hello to Gemini 3! SOTA reasoning & multimodal understanding, and our best model for coding :)
-
@emostaque
Emad
on x
The most interesting thing testing Gemini 3 Pro has been how *efficient* it is from tokens to tool calls The intelligence per token of models is increasingly rapidly even as prices fall, its quite something
-
@cursor_ai
@cursor_ai
on x
Gemini 3 Pro is now available in Cursor!
-
@jeffdean
Jeff Dean
on x
I'm really excited about our release of Gemini 3 today, the result of hard work by many, many people in the Gemini team and all across Google! 🎊 We've built many exciting new product experiences with it, as you'll see today and in the coming weeks and months. You can find it [ima…
-
@benbajarin
Ben Bajarin
on x
Max has been analyzing Gemini 3 early, check out his thoughts on what looks like sustained model leadership from @Google TPUs!
-
@skirano
Pietro Schirano
on x
I asked Gemini 3 Pro to create a 3D LEGO editor. In one shot it nailed the UI, complex spatial logic, and all the functionality. We're entering a new era. https://t.co/...
-
@mattshumer_
Matt Shumer
on x
I've had access to Gemini 3 since November 13th. Since then, I've used it as my daily-driver, pushing it to its limits. Here's my review of Gemini 3: https://shumer.dev/...
-
@thefox
Nick Fox
on x
Gemini 3 is available *starting now* in AI Mode in Google Search!! For the first time, our latest Gemini model is shipping in Search on day one. From model to product, faster than ever... 🚀
-
@scaling01
@scaling01
on x
Gemini 3 takes first place on the Artificial Intelligence Index [image]
-
@googleai
@googleai
on x
Today we're taking a big step on the path toward AGI and releasing Gemini 3— our most intelligent model yet. With Gemini 3, you can bring any idea to life. It is state-of-the-art in reasoning, the best model in the world for multimodal understanding, and our best agentic and [vid…
-
@googleai
@googleai
on x
The first model in the Gemini 3 family is Gemini 3 Pro, which is rolling out globally TODAY. We're bringing Gemini 3 to more products than we ever have before. Here are a few ways to try it: [image]
-
@sundarpichai
Sundar Pichai
on x
Introducing Gemini 3 ✨ It's the best model in the world for multimodal understanding, and our most powerful agentic + vibe coding model yet. Gemini 3 can bring any idea to life, quickly grasping context and intent so you can get what you need with less prompting. Find Gemini [vid…
-
@google
@google
on x
Introducing Gemini 3 — our most intelligent model that helps you bring any idea to life. Gemini 3 is our next step on the path toward AGI and has: 🧠 State-of-the-art reasoning 🖼️ Deep multimodal understanding 💻 Powerful vibe coding so you can go from prompt to app in one shot [im…
-
@bilawalsidhu
Bilawal Sidhu
on x
Hot off the presses is Gemini 3 Pro, Google's new SOTA model - tops LMArena with a score of 1501 points. Launching simultaneously as an API, inside the consumer Gemini app, Google Search, oh and a new agentic IDE. Here's the TL;DR: 1. Gemini 3 Pro: new SOTA for multimodality [vid…
-
@altryne
Alex Volkov
on x
Gemini was always a beast at multimodal, but 3 takes it even further! Earning impressive marks of 81% on MMMU-Pro and 87.6% on Video-MMMU, it achieves a leading 72.1% on SimpleQA Verified, demonstrating significant strides in factual accuracy. Text, Image and Video as inputs! [im…
-
@altryne
Alex Volkov
on x
IT'S HERE! The new state of the art AI intelligence in the form of Gemini 3-pro - and it's available in Search AI mode, API, @GoogleAIStudio and Vertex today! Plus: 🧩 Generative UIs! 🤖 Agent mode 🧑💻 VSCode clone 🤔 3 Deep Think And lots more! Let's break it down 👇 [video]
-
@altryne
Alex Volkov
on x
Evals are nuts 📈 With a massive 1501 @arena (+50 on 2.5 pro), 91.9% on GPQA diamond and 37.5% on HLE, this is the most powerful LLM intelligence we've had access to yet. And w/ DeepThink + tools an even more impressive 41.1% on HLE and an unprecedented 45.1% on @arcprize [video]
-
@scaling01
@scaling01
on x
Gemini 3 Pro going absolutely vertical on Vending Bench [image]
-
@theo
@theo
on x
Google just dropped Antigravity, their own VS Code fork. It includes Gemini 3 Pro for free. I got early access. Just dropped my video about it! [image]
-
@googledeepmind
@googledeepmind
on x
Google Antigravity is our new agentic development platform. It helps developers build faster by collaborating with AI agents that can autonomously operate across the editor, terminal, and browser. It uses Gemini 3 Pro 🧠 to reason about problems, Gemini 2.5 Computer Use 💻 for [vid…
-
@steren
@steren
on x
I've been using Google @antigravity for a few days, here are my impressions: To me, it's not an IDE, it's a coding agent UI, powered by Gemini 3 Pro. Notably, the agent can control a Chrome browser, enabling it to validate open and use the app and it builds. [image]
-
@antigravity
@antigravity
on x
Meet Google Antigravity, your new agentic development platform. An evolution of the IDE, it's built to help you: - Orchestrate agents operating at a higher, task-oriented level - Run parallel tasks with agents across workspaces - Build anything with Gemini 3 Pro. [video]
-
@_mohansolo
Varun Mohan
on x
Excited to launch Google Antigravity, our next generation agentic IDE, now powered by Gemini 3! [video]
-
@mweinbach
Max Weinbach
on x
I've had access to Google's new Antigravity IDE and Gemini 3 Pro for the past few days and they are really good. Gemini 3 Pro is prob my favorite agentic model, but needs more prompting than GPT-5.1, and Antigravity has replaced Cursor for me (for now) https://creativestrategies.…
-
@rohanlikesai
Rohan Doshi
on x
🚀 We just launched Gemini 3 Pro — the strongest multimodal understanding model ever built. I lead product for Gemini's multimodal vision capabilities, and I want to share more about the massive wins we are seeing across document, screen, spatial, and video understanding. 🧵
-
@tengyanai
@tengyanai
on x
Looks like x'mas came early: Gemini 3 is here! i don't put much weight on benchmarks anymore (prefer to vibe test them for my tasks), but... it looks v good beats GPT-5.1 on almost everything.
-
@mattshumer_
Matt Shumer
on x
Yeah, AI progress is totally, definitely stalling... Look at MathArena Apex. GPT-5.1 scored 1%. Gemini 3 scored 23%. That is a >20x jump on one of the hardest reasoning tasks we have. But sure, keep your head in the sand... [image]
-
@sama
Sam Altman
on x
Congrats to Google on Gemini 3! Looks like a great model.
-
@ludwigabap
Ludwig
on x
re: Gemini 3.0 Benchmarks do NOT matter and 99% of them are awful and built by idiots. There is a trillion dollar marketing war being waged for your mind and time. Everyone is paid for. Nothing matters until it does and nothing is real until you can use it.
-
@jeffdean
Jeff Dean
on x
Gemini 3 also scores well on the lmarena leaderboards, ranking #1 across all the major @arena leaderboards.
-
@arena
@arena
on x
🚨BREAKING: @GoogleDeepMind's Gemini-3-Pro is now #1 across all major Arena leaderboards 🥇#1 in Text, Vision, and WebDev - surpassing Grok-4.1, Claude-4.5, and GPT-5 🥇#1 in Coding, Math, Creative Writing, Long Queries, and nearly all occupational leaderboards. Massive gains [image…
-
@officiallogank
Logan Kilpatrick
on x
Gemini 3 Pro with the largest delta recorded thus far on @Designarena 🤯 [image]
-
@mikeknoop
Mike Knoop
on x
We just verified Gemini 3 Pro and Deep Think (Preview) are over 2X SOTA on ARC v2! This is really impressive and frankly a bit surprising. Impressive because many of the v2 solves indicate clear complexity scaling over v1. Such as tasks 65b59efc, e3721c99, and dd6b8c4b We're
-
@levie
Aaron Levie
on x
At Box, we've been testing Gemini 3 Pro in early access with Box AI on our most complex advanced reasoning eval, and Gemini Pro was a massive 22 percentage point improvement over Gemini 2.5 Pro. For this test, we ask the model a series of complex, real-world questions with a set …
-
@artificialanlys
@artificialanlys
on x
For the first time, Google has the most intelligent model, with Gemini 3 Pro Preview improving on the previous most intelligent model, OpenAI's GPT-5.1 (high), by 3 points [image]
-
@artificialanlys
@artificialanlys
on x
As well as a substantial improvement in intelligence, Gemini 3 Pro Preview has demonstrates increased token efficiency compared to Gemini 2.5 Pro, using fewer tokens on the Artificial Analysis Intelligence Index than its predecessor, as well as other leading models such as Kimi […
-
@artificialanlys
@artificialanlys
on x
Gemini 3 Pro Preview leads two of the three coding evaluations in the Artificial Analysis Intelligence Index, including an impressive 56% in SciCode, an improvement of over 10 percentage points from the previous highest score. It is also strong in agentic contexts, achieving the …
-
@artificialanlys
@artificialanlys
on x
Gemini 3 Pro Preview takes the top spot on the Artificial Analysis Omniscience Index, our new benchmark for measuring knowledge and hallucination across domains. Gemini 3 Pro Preview comes in first for both Omniscience Index (our lead metric that takes off points for incorrect [i…
-
@artificialanlys
@artificialanlys
on x
Individual results across the 10 evals we run independently for the Artificial Analysis Intelligence Index: MMLU-Pro, GPQA Diamond, Humanity's Last Exam, LiveCodeBench, SciCode, AIME 2025, IFBench, AA-LCR, Terminal-Bench Hard, 𝜏²-Bench Telecom [image]
-
@kimmonismus
@kimmonismus
on x
tough times ahead for OpenAI, Anthropic and xAI. And I have a feeling they won't lose their lead anytime soon. The power of the sheer number of TPUs.
-
@scaling01
@scaling01
on x
Its so over for OpenAI and Anthropic. Gemini 3 Pro Benchmarks 37.5% on HLE 31.1% on ARC-AGI-2 2439 Elon on LiveCodeBench Pro 85.4% on Tau-Bench 72.1% on SimpleQA Verified SOTA everywhere except SWE-Bench Verified [image]
-
@dillonuzar
Dillon Uzar
on x
Context Arena Update: Added Gemini 3.0 Pro Preview (Thinking, 11-18) to the MRCR leaderboards. It establishes a new state-of-the-art in context performance, taking the #1 spot on all our AUC leaderboards and for nearly all pointwise scores. All results at: [image]
-
@ryanlpeterman
Ryan Peterman
on x
Gemini Pro 3 benchmark result I'm most excited about, would love to never have to click through random booking UIs ever again (e.g. reservations, flights, etc) [image]
-
@kimmonismus
@kimmonismus
on x
Gemini 3 is a beast of a model. It set a new benchmark for LLMs [image]
-
@alexfinn
Alex Finn
on x
Gemini 3 benchmarks leaked...wow Highest in every benchmark except coding (just behind Sonnet 4.5) Truly mind blowing If your boss doesn't give you the afternoon off to play with super intelligence, you 100% should quit your job now [image]
-
@chiyeung_law
Ziyang Luo
on x
🎉 Huge congrats to @GoogleDeepMind on the release of Gemini-3.0-Pro, which scored an impressive 72.7 on our ScreenSpot-Pro benchmark, outperforming specialized GUI-focused models by over 6 points. 🚀 Also excited to see ScreenSpot-Pro now widely adopted by almost all leading [imag…
-
@deryatr_
Derya Unutmaz
on x
Gemini 3.0 Pro absolutely dominates every benchmark! The jump from 2.5 is nuts! Its scores on the most difficult benchmarks suggest this is essentially baby AGI! Humanity Last exam: 37.5% ARC-AGI-2: 31.1% LiveCodeBench Pro: 2439 Math arena apex : 23.4% Simple QA: 72.1% [image]
-
@__apf__
Adriana Porter Felt
on x
mom MOM look sota on all benchmarks except swe bench very good dear there's more soda in the fridge if you want some [image]
-
@chasebrowe32432
Chase Brower
on x
Gemini 3 Pro (preview) scores 91% on VPCT (spatial reasoning) Uhhhh jesus christ [image]
-
@officiallogank
Logan Kilpatrick
on x
And say hello to Gemini 3 Deep Think, even more SOTA compared to Gemini 3 Pro 🤯 [image]
-
@officiallogank
Logan Kilpatrick
on x
Introducing Gemini 3 Pro, the world's most intelligent model that can help you being anything to life. It is state of the art across most benchmarks, but really comes to life across our products (AI Studio, the Gemini API, Gemini App, etc) 🤯 [image]
-
@saboo_shubham_
Shubham Saboo
on x
Gemini 3 pro one-shotted an SVG of a pelican riding a bicycle. [video]
-
@saboo_shubham_
Shubham Saboo
on x
Introducing Gemini 3 Pro > best model in world for multimodal understanding > SOTA at complex reasoning > insanely good at agentic tool use and vibe-coding > great at zero-shot code generation > beats Claude sonnet 4.5, GPT-5.1 on almost every benchmark [image]
-
@artificialanlys
@artificialanlys
on x
Gemini 3 Pro is the new leader in AI. Google has the leading language model for the first time, with Gemini 3 Pro debuting +3 points above GPT-5.1 in our Artificial Analysis Intelligence Index @GoogleDeepMind gave us pre-release access to Gemini 3 Pro Preview. The model [image]
-
@taumuyi
Tau-Mu Yi
on bluesky
You can look at the benchmarks and Gemini 3 is an incremental advance (like GPT-5), but it is a significant incremental advance across benchmarks #AI [embedded post]
-
@googleresearch
@googleresearch
on x
This Generative UI capability is built on Google's advanced foundation models & pioneering research from Google Research. This work built the core principles & first working prototype behind generative UI that unlocks the new interactive tools in both Search & the Gemini app. 📃
-
@googleresearch
@googleresearch
on x
AI Mode in Search dynamically creates the ideal visual layout for responses, tailored to your query. When the model detects that an interactive tool will help you better understand the topic, it uses its generative UI capabilities to code a custom simulation or tool in real time …
-
@googleresearch
@googleresearch
on x
In dynamic view, Gemini designs and codes a custom user interface in real-time, perfectly suited to your prompt. In our evaluation, raters strongly preferred dynamic view's personalized assets over standard LLM output. [video]
-
@googleresearch
@googleresearch
on x
Generative UI is a capability where an AI model generates dynamic, visual layouts & interactive interfaces, such as web pages, games, & apps. Today with the launch of Gemini 3, you can experience our novel implementation of generative UI in @GeminiApp & Google Search. 🧵↓ & blog […
-
@jeffdean
Jeff Dean
on x
As an example of how we are building on top of Gemini 3, AI Mode in Search now uses Gemini 3 to enable new generative UI experiences, all generated completely on the fly based on your query. Here's how you might use this to learn a complex topic like how RNA polymerase works. [vi…
-
@rajanpatel
Rajan Patel
on x
Say you're a student learning about RNA polymerase. AI Mode's new generative UI can create dynamic layouts with interactive tools and simulations to aid in deeper comprehension so you can learn even more interactively while exploring content across the web. [video]