Moonshot AI releases Kimi K3, a 2.8T-parameter AI model that it says rivals Claude Opus 4.8 and GPT-5.5, and plans to release its full model weights by July 27
Today, we are introducing Kimi K3 — our most capable model. Kimi K3 is a 2.8T-parameter model built on our Kimi Delta Attention and Attention Residuals …
Kimi
Context & Ripple Effects
Moonshot’s Kimi line has moved from K2’s mixture-of-experts architecture and benchmark claims to K2.5’s agent-oriented capabilities and K2.6’s open-weight release for long-horizon coding. K3 extends that sequence with a substantially larger model and another planned weight release.
The significance is not only Moonshot’s claimed frontier performance, which remains a vendor claim, but its continued use of open weights as part of the Kimi product strategy. That puts deployability alongside raw model capability in the competitive comparison.
First-order effects
- Moonshot will make Kimi K3’s weights available on its stated timetable, giving developers and enterprises a new model to evaluate and potentially run or adapt outside a purely hosted API.
- Kimi K3 immediately raises the competitive bar Moonshot is claiming against Opus 4.8 and GPT-5.5, while shifting attention to independent testing of those performance assertions.
Second-order effects
- Open-weight availability gives model buyers more leverage: they can compare proprietary frontier offerings with an alternative they may be able to customize and deploy under their own operating constraints.
- Competing model providers face added pressure to differentiate on reliable task performance, tooling, deployment options, and economics—not solely on headline benchmark positioning.
Third-order effects
- If increasingly capable models are released with usable weights, frontier-model competition may become less concentrated around closed API access and more centered on distribution, integration, and the cost of delivering useful work.
- The pattern also makes infrastructure a sharper dividing line: broader access to large model weights can expand experimentation, but practical deployment will remain shaped by the compute needed to serve and adapt them.
The trend: K3 is part of the industrialization of AI models, in which frontier-capability claims are increasingly paired with open-weight distribution and competition shifts toward deployment economics and product reach.
Related: AI industrialization · Model buyer power · Compute-capacity economics · Moonshot AI · Moonshot introduces Kimi K2.6, an open-weight model that it says shows · Moonshot's Kimi K2 uses a 1T-parameter MoE architecture with 32B activ
Related Coverage
- China's open-weight Kimi model stuns AI world with frontier-level results Axios · Madison Mills
- China's Moonshot AI chases ‘DeepSeek moment’ with much-hyped model The Economic Times
- China's Moonshot AI Releases Model to Challenge Top U.S. Systems Wall Street Journal
- China's Moonshot launches world's largest open AI model, nears US rivals Business Standard
- China Just Dropped Another Bomb on America's Frontier AI Companies Gizmodo · Ece Yildirim
- Kimi K3, and what we can still learn from the pelican benchmark Simon Willison's Weblog · Simon Willison
- China's Moonshot AI Set To Narrow Anthropic Performance Gap Silicon UK · Matthew Broersma
- Early look at Kimi K3 generations from Moonshot AI on Arena TestingCatalog AI News · Alexey Shabanov
- Kimi K3 shows open AI models have finally caught up with proprietary US-based rivals Neowin · Pradeep Viswanathan
- China's Moonshot throws down the gauntlet with Kimi K3, the world's largest open-weights model SiliconANGLE · Mike Wheatley
- Moonshot's Kimi K3 pushes Chinese AI into Fable-level territory Fortune · Nicholas Gordon
- Moonshot AI Releases Kimi K3: A 2.8 Trillion Parameter Open MoE Model With Kimi Delta Attention and 1M Context MarkTechPost · Asif Razzaq
- Kimi K3: 2.8T Parameter Model Ships With 1M Context Blockchain.News · Soumith Chintala
- Kimi's open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AI The Decoder · Matthias Bastian
- Moonshot AI launches Kimi K3 Constellation Research · Larry Dignan
- Moonshot Launches Kimi K3 With 2.8 Trillion Parameters and 1M Context Implicator.ai · Marcus Schuler
- Kimi K3, a Chinese open weight model from Moonshot AI, is a frontier class model that is going toe to toe with Claude Fable and GPT 5.6. So much for Chinese models being months behind American models. They've already caught up at a fraction of the token price. Paying Fable prices seems silly now. @carnage4life · Dare Obasanjo
- Kimi K3: Open Frontier Intelligence Hacker News
- Kimi K3 Intelligence, Performance & Price Analysis Artificial Analysis
- Moonshot AI's New Kimi K3 Challenges U.S. Frontier Models The Information · Juro Osawa
- Chinese startup launches “Kimi K3,” the world's largest open-weights AI model, approaching the performance of Fable and GPT-5.6 The Hans India · Kahekashan
- What is Kimi K3? China's new AI model aces Claude Fable 5 and GPT-5.6 Financial Express · Amritanshu Mukherjee
- Moonshot AI unveils world's largest open-source AI model as China narrows gap with US rivals South China Morning Post
- What is Kimi K3: Moonshot's frontier AI model that matches Claude Opus 4.8 and GPT 5.5 Digit · Vyom Ramani
- [AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing Latent.Space
- Research|LLM: Kimi K3 - Scaling Still Works; An Expensive Model Competing at Front Tier FUNDA
- Everything That Happened in AI Today (Thursday, July 16, 2026) The Neuron · Grant Harvey
- Chinese AI start-up Moonshot launches model challenging Anthropic's lead Financial Times
- What is Kimi K3? Why this new Chinese AI model is making headlines Business Today
- China just erased America's AI lead Axios
- Moonshot AI unveils Kimi K3, the world's largest open-weight AI model: What to know The Indian Express
- Moonshot AI Unveils Kimi K3: A 2.8 Trillion-Parameter Model Challenging US AI Giants Blockonomi · Trader Edge
- New open-weight AI from China is toppling the best of OpenAI and Claude Fable Digital Trends · Sudhanshu Kumar Mangalam
- China's Kimi K3 AI Model May Trigger a DeepSeek Moment. Expect a Lot of FUD. Key Context · Tae Kim
- Decoded: What is Kimi K3, an open AI model challenging OpenAI and Anthropic Business Standard · Aashish Kumar Shrivastava
- China's Moonshot unveils world's largest open AI model, closing in on US rivals Reuters · Laurie Chen
- Moonshot AI Unveils 2.8T-Parameter Kimi K3 AI Model WinBuzzer · Markus Kasanmascheff
- Moonshot's Kimi K3 closes the frontier gap The Rundown AI
- China's Moonshot launches Kimi K3 open AI model Verdict · Anwesha Pattanaik
- China's Powerful New AI Surprises Investors, Fueling Tech Rout Bloomberg
- China's 2.8-trillion-parameter Kimi K3 beats Claude Fable 5 in Frontend Code Arena benchmark— Moonshot AI delivers largest open-weight AI model ever, as China works around U.S. compute limits Tom's Hardware · Luke James
- Moonshot unveils Kimi K3, largest open-weight AI model yet Silicon Republic · Suhasini Srinivasaragavan
- 😼 Kimi K3 goes open The Neuron · Grant Harvey
- How China Is Taking Control of the Future of AI Bloomberg · Catherine Thorbecke
- MoonshotAI: Kimi K3 moonshotai/kimi-k3 OpenRouter
- Chinese startup Moonshot AI unveils Kimi model it says rivals OpenAI, Anthropic CNBC · Kai Nicol-Schwarz
- China's Powerful New AI Surprises Investors, Fueling Tech Rout Bloomberg
- US stock futures, Asian markets down on concerns over Chinese AI advances CNN · Chris Isidore
- Moonshot AI Kimi K3 launch sends rival AI stocks lower Quartz · Cris Tolomia
- Kimi K3 Built A Chip In Just 48 Hours, Which Pushes Over 8700 Tokens/s, As China's Moonshot Delivers A 2.8 Trillion Parameter Frontier AI Model Wccftech · Hassan Mujtaba
- Bitcoin faces fresh headwinds as China's Kimi beats Claude, GPT in coding benchmark CoinDesk · Shaurya Malwa
- A New Blow to Tracking Gun Sales New York Times
- China's Moonshot AI launched its largest model ever — and rival AI stocks plunged Quartz · Cris Tolomia
- Aaron Levie's Post LinkedIn · Aaron Levie
- Kimi (Moonshot AI)'s Post LinkedIn
- Futures Tumble As Latest Chinese “DeepSeek Moment” Sparks Chip Meltdown ZeroHedge News · Tyler Durden
- 3 reasons China's Kimi K3 is turning heads in Silicon Valley Business Insider · Thibault Spirlet
- The hype around Kimi K3 should not alarm Washington, because while K3 is very good, it is not frontier-level and likely lacks dangerous cyber capabilities Transformer
- Clouded Judgement 7.17.26 - Open Weights, Closed Prices? Clouded Judgement · Jamin Ball
- Kimi K3: Open Frontier Intelligence Hacker News
- Moonshot unveils Kimi K3, largest open-weight AI model yet Silicon Republic · Suhasini Srinivasaragavan
- China's Powerful New AI Surprises Investors, Fueling Tech Rout Bloomberg
- China Just Unveiled An AI Model That Rivals America's Best. Silicon Valley Is No Longer Clearly Leading The Race. International Business Times · Merin Rebecca Thomas
- China's Kimi K3 Identifies Itself As Anthropic's Claude In At Least One Conversation, Betraying Its Distilled Origins Wccftech · Rohail Saleem
- What smart people are saying about China's hot new Kimi K3 AI model Business Insider
- Kimi K3's performance bolsters Sacks' case against AI regulation Axios · Madison Mills
- Report: China's ‘Kimi K3’ AI Model Matches Performance of Top American Systems Breitbart · Lucas Nolan
- Chinese AI firm Moonshot unveils powerful model with capabilities close to Anthropic, OpenAI New York Post · Thomas Barrabi
- China's Moonshot AI Unveils Kimi K3, a Massive Parameter Model Challenging U.S. Dominance Techstrong.ai · Jon Swartz
- Moonshot reveals new AI model, and it's a big surprise — here's why Kimi K3 is a threat to the likes of OpenAI TechRadar · Darren Allan
- Nasdaq Drops 2% As Global Investors Panicked Over New Chinese AI Advances, Moonshot AI Ventureburn · Ekemini
- Kimi K3, and what we can still learn from the pelican benchmark Hacker News
- Kimi K3 tops Arena's coding leaderboard — and it's open-weight The New Stack · Amanda Caswell
- Moonshot's Kimi K3 Won't Fit on a Single Nvidia DGX B200 Implicator.ai · Marcus Schuler
- Former White House Crypto Czar David Sacks Warns US Could Lose AI Race CoinGape
- New Chinese AI chatbot Kimi K3 rivals US leaders in the field Washington Examiner · David Zimmermann
- Moonshot AI's Latest Model Dings Tech Stocks. We've Seen This Story Before. Barron's Online · Nate Wolf
- Chinese AI Startup Moonshot Unveils Kimi K3 Model—Will It Challenge OpenAI And Anthropic? Forbes · Ty Roush
- Kimi K3 threatens AI business models Semafor · Reed Albergotti
- How to Work Effectively with GPT-5.6 Towards Data Science · Eivind Kjosbakken
- Moonshot debuts 2.8 trillion-parameter Kimi K3 Tech in Asia · Naomi Li Gan
Discussion
-
@zijing_wu
Zijing Wu
on x
Chinese AI start-up Moonshot to launch model challenging Anthropic's lead * Set to release as early as tonight * 2-3T, largest Chinese model to date * Benchmark performance Opus 4.8 < K3 < Fable * Attention Residuals & Kimi Linear * Fundraising at $31.5B https://as.ft.com/...
-
@kimi_moonshot
@kimi_moonshot
on x
Introducing Kimi K3: Open Frontier Intelligence 🔹 2.8 Trillion Parameters, 1 Million Context, Native Multimodal 🔹 Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts 🔹 Attention Residuals deliver ~25% higher training efficiency at <2% additional [ima…
-
@theahmadosman
Ahmad
on x
Kimi K3 - 2.8T Parameters - 1M Context Length Benchmarks > GDPval-AA v2: 3rd place, ranks > directly below GPT 5.6 Sol Max & Fable 5 Max > AA-Briefcase: 2nd place, beats GPT 5.6 Sol Max, right below Fable 5 Max > BrowseComp: 1st place, beating GPT 5.6 Sol Max & Fable 5 Max [image…
-
@sriramk
Sriram Krishnan
on x
kimi k3 is a big moment with multiple implications for the entire industry.
-
@kimi_moonshot
@kimi_moonshot
on x
Meet Kimi K3 [video]
-
@tszzl
Roon
on x
the era of the chinese labs being far behind is over, Kimi is at least on par with the modern public frontier models. people have to think differently now without any competitive margin built in
-
@tszzl
Roon
on x
the world vision of open weights models running themselves, self replicating, training new versions of themselves (at least the kind of behavioral modifications that won't require massive compute scale), is really not very far away
-
@eliebakouch
Elie
on x
Kimi K3 (2.8T total parameter) will be the biggest open weight model ever and by far grok 4.5 is 1.5T just for comparison, deepseek v4 is 1.6T, opus and gpt 5.5 are said to be around 1.5T
-
@yuchenj_uw
Yuchen Jin
on x
> Kimi K3 ranks only behind Claude Fable 5 Max and GPT-5.6 Sol Max, and surpasses Claude Opus 4.8 Max's score of 1600. I'm super excited for Kimi K3 to be open-sourced! [image]
-
@cline
@cline
on x
Moonshot's new Kimi K3 is the largest open weight model ever at 2.8T param, and scores the same as GPT-5.6 and Fable 5 on benchmarks. However, it is signficantly cheaper at $3/M input, $15/M output - same price as Sonnet 5. This is a game changing milestone for open weights. [ima…
-
@ryangreenblatt
Ryan Greenblatt
on x
Kimi K3 was significantly but not massively above my expectations. I'd tentatively guess it's similar in overall usefulness/usability to Opus 4.8 and in overall capability somewhat above Opus 4.8 (while also being somewhat more benchmaxxed). As a pretrain, it's probably somewhere
-
@kimmonismus
@kimmonismus
on x
Official In benchmarks, Kimi k3 is reportedly just behind GPT-5.6 and Fable 5, but ahead of Opus 4.8. However, in terms of price, it's on par with Sonnet 5. An absolute game changer. Insane! [image]
-
@scaling01
@scaling01
on x
what's pretty clear, even without knowing all of these other benchmarks is that China now has models that are genuinely useful for speeding up their own model development
-
@emostaque
Emad
on x
US labs gonna end up distilling Chinese models
-
@adonis_singh
Adi
on x
kimi-k3 & muse-spark-1.1 on eyebench! muse-spark is actually doing really good, topping out all models except gpt-5.4-5.6 kimi scores 4 points lower & costed ~10x more than muse-spark... not a good look. [image]
-
@emollick
Ethan Mollick
on x
Kimi K3 seems really good, closest to the frontier yet, but also wow does the model/harness love to loop back over and over again over tasks tweaking and changing things at max level. Anyhow, examples incoming, when they finish.
-
@rauchg
Guillermo Rauch
on x
Kimi K3 is the best performing model on https://nextjs.org/evals, ahead of Fable, reaching a comparable success rate in less time. This is the first time that an open model is ahead of all proprietary ones for this comprehensive web engineering benchmark. Notes: ▪️ Benchmarks [im…
-
@scaling01
@scaling01
on x
Kimi-K3 is 10th in the Text Arena [image]
-
@scaling01
@scaling01
on x
The moonshot was successful the small chinese lab is taking down trillion dollar companies
-
@natolambert
Nathan Lambert
on x
@arena @Kimi_Moonshot at this point the distillation arguments need to die and understand that China is also very good at building models
-
@emostaque
Emad
on x
something something recursive self improvement [image]
-
@kimi_moonshot
@kimi_moonshot
on x
Internal knowledge work bench Beyond public benchmarks, Kimi K3 Max also shows consistent gains on our internal benchmarks, which are built from recurring patterns and challenges in real-world user-agent workflows. It scores 75.5 on Online Exp Bench, 73.5 on DECK-Bench, and [imag…
-
@aravsrinivas
Aravind Srinivas
on x
Incredible
-
@thegenioo
Hamza
on x
I was wrong about Kimi K3 Just tried it in Claude Code and it is Fable 5 level for sure Far better than GPT-5.6 Sol at least at front-end It is really slow and consumes a lot of tokens but it is so good at completing its task, I like what I see here @Kimi_Moonshot [video]
-
@kimi_moonshot
@kimi_moonshot
on x
K3 is built on Kimi Delta Attention (KDA) and Attention Residuals (AttnRes), two architectural updates designed to improve how information flows across sequence length and model depth. We have also scaled up Mixture of Experts (MoE) sparsity, effectively activating 16 out of [ima…
-
@jukan05
Jukan
on x
My conclusion so far is that Kimi K3 may be the first model to narrow the gap with leading U.S. closed-source models to less than three months.
-
@theojaffee
Theo Jaffee
on x
Registering my prediction of no widespread societal chaos after the open-sourcing of Kimi K3
-
@_maxblade
Max Blade
on x
KIMI K3 is here, and it blew my mind 🤯 yes, it is expensive and very slow Took over 30 minutes, and blew an ENTIRE $20 kimi sub before it even finished the one prompt for this game. BUT OMG... the output geniuinly has fable 5 level magic built into it, this model falls [video]
-
@laschuk
@laschuk
on x
I put Kimi K3 up against Fable 5 in my benchmark. The test: clone Apple's homepage. Single prompt, one shot, zero help. Kimi K3: $0.44 (at left) Fable 5: $0.94 (at right) Same test. Same prompt. Half the price. Fable runs $10/$50 per MTok, the most expensive tokens on the [video]
-
@jeffwang
Jeff Wang
on x
Chinese open source is no longer “6 months behind”, but it's also no longer “10% of the cost” either Congrats to Moonshot on Kimi K3, 2.8T, $3/$15 - same as Sonnet but initial feedback is great, we'll be sure to run FrontierCode on our side
-
@ryangreenblatt
Ryan Greenblatt
on x
I did some quick tests that indicated that the Kimi K3 pretrain is around halfway between Opus 4 and Opus 4.5. So ~10 months behind Anthropic. These tests probably understate data improvements, so overall I think it's a similarly good pretrain to Opus 4.5 (~8 months behind).
-
@zephyr_z9
@zephyr_z9
on x
Crazy big boi numbers from Kimi K3 [image]
-
@yzhang_cs
Yu Zhang
on x
In the early morning, @nathancgy4, @Xinyu2ML, @Yulun_Du and I were preparing some showcases for the blog while watching the World Cup. The moment Argentina beat England, I felt something. I looked up, and saw the most unforgettable Beijing sunrise. So I took the photo. I knew [im…
-
@yacinemtb
Kache
on x
insane. they distilled american frontier models that haven't even been invented yet. damn chinese
-
@dhh
@dhh
on x
I just hoped back on Kimi K2.7 for some work yesterday. Hadn't been using it since GPT 5.6 and Fable. But great reminder of just how good these open-weight models have gotten. For what I was doing, Kimi was faster and just as good. Excited for K3!
-
@emostaque
Emad
on x
Kimi K3 by @Kimi_Moonshot is a true frontier model, open source! Congrats! We don't have all details (training tokens etc), but would estimate fp16 pretraining w/ muon => MXFP4/MXFP8 for SFT stage ~1e25 flops total (~Inkling!) Total cost: $15-$25m Scale isn't all you need! [image…
-
@mweinbach
Max Weinbach
on x
Ok it's up TLDR Kimi K3 is fine but there are cheaper and better models Muse Spark 1.1 and Kimi K2.6 should not be considered at all. https://csbench.com/...
-
@deredleritt3r
Prinz
on x
A few thoughts on Kimi K3: - The Moonshot team is absolutely cracked. Some of you might remember that K1.5 was released on January 20, 2025 - i.e., on the same day as DeepSeek R1, and it was ~just as good as R1. IIRC, the model wasn't open-source, which is why no one ever
-
@nrehiew_
@nrehiew_
on x
Here are a few benchmark scores of K3 that have been officially confirmed This is a Fable/Sol class model that is strictly better than Opus 4.8 across the board at Sonnet pricing. Insane [image]
-
@theahmadosman
Ahmad
on x
Kimi K3: The Opensource AI Frontier [image]
-
@wongmjane
Jane Manchun Wong
on x
Unfortunately, Kimi K3 is not even willing to discuss 9/11 [image]
-
@mweinbach
Max Weinbach
on x
At $3/$15 in out tokens eh, like i get why it's priced what it is and the compute needed but i have no desire to pay that given it doesn't seem token efficient
-
@bridgemindai
@bridgemindai
on x
Kimi K3 just beat Fable 5 on the BridgeBench Horror House game test. I did not expect this at all. Kimi K3 is better than Fable 5 at game development and UI design. The two things Fable 5 was supposed to own. The 5x price increase suddenly makes sense. [video]
-
@mweinbach
Max Weinbach
on x
Kimi K3 being 2.8T parameter and 1M context is cool but show me the sparsity, show me the price How quick can the inference providers scale this to 200 tok/s This is what I care about!!!! Efficient HUGE models
-
@thegenioo
Hamza
on x
Kimi K3 is out. But it's not what I expected, not better than Fable or Opus 5 level. It is really good, persistent and determined, but tokens hungry and also very very slow. The designs are not out of this world on my test, but very clean, accurate and one shot. [video]
-
@oluwaphilemon1
@oluwaphilemon1
on x
Kimi K3 is currently leading Claude Fable 5, GPT-5.6, and Grok 4.5 creating game demo. A mystery Arena .ai model called “Kivine” is reportedly Kimi K3. In this AI Game Watch, I break down the first game and 3D demos shared by creators on X. [video]
-
@norwakar
Diwakar Ray Yadav
on x
Kimi K3 just went global and beta testers are saying it matches Opus-level output at nearly half the price. 🚀 China clearly isn't slowing down in this AI race. Overhyped or the real deal? Drop your take 👇 [image]
-
@kimmonismus
@kimmonismus
on x
Kimi k3 starts rolling out. Official release is imminent! Super freaking excited for the evals. Will it bear opus 4.8? Could be a real game changer, literally.
-
@yuhasbeentaken
@yuhasbeentaken
on x
They distilled Fable, now they're distilling Anthropic's marketing. That is a joke about Kimi K3. But,it's working. Copy the frontier, undercut the price, ship faster than the lab you copied from. If K3 lands anywhere near the rumors, expect this to keep happening. Not because [i…
-
@zackkorman
Zack Korman
on x
If Kimi 3 is actually better than Opus 4.8 then I just want to say it was good knowing you all. Stay strong through the cyber apocalypse.
-
@theo
@theo
on x
Normally I don't comment on rumors, but if Kimi K3 actually beats out Opus 4.8 that's nuts Also hearing Opus 5 might drop?
-
@chetaslua
@chetaslua
on x
Kimi K3 vs GPT-5.6 Sol Difference in taste is so stark , like if i swap kimi k 3 name with fable 5 people will trust it but this is not only about visuals , its also about function , you can see both achieve same result but the way to achieve is different kimi is more [video]
-
@atomtanstudio
@atomtanstudio
on x
For everyone getting excited about Kimi K3. They are not cheap. [image]
-
@kimmonismus
@kimmonismus
on x
Kimi k3 is being released tonight, via FT -2-3t parameters (Opus4.8 has about 1.5t) -1m context -Expected to exceed Opus 4.8 performance! The time when China was six months behind is over. History is presumably being made today. [image]
-
@jun_song
Jun Song
on x
Kimi-K3 vs Opus-4.8 Flappybird test. Kimi is significantly better than Opus. That's the reason why I claimed it as Opus-5 level. Test done in @arena [video]
-
@synthwavedd
Leo
on x
I think Kimi K3 is going to shock some of the “the Chinese are 8 months behind the Western frontier” people
-
@shanumathew93
Shanu Mathew
on x
Reminder that Opus 4.8 came out at thee end of May. So frontier leadership went from 6+ months to 1-1.5 months. Unclear how much of this is because of wide distillation and how far internal models are at American labs but the direction is clear that China is catching up...
-
@sungkim
Sung Kim
on bluesky
🔹 Built for long-horizon agentic coding and self-evolving workflows — Tech blog: kimi.com/blog/kimi-k3 [image]
-
r/LocalLLaMA
r
on reddit
Kimi K3: Open Frontier Intelligence
-
@artificialanlys
@artificialanlys
on x
Kimi K3 scores 57 on the Artificial Analysis Intelligence Index. Its intelligence is comparable to Opus 4.8 and GPT-5.5 but remains behind Fable 5 and GPT-5.6 Sol. Moonshot AI has expressed plans to release the 2.8T parameter model's weights, which would make it the leading open …
-
@arena
@arena
on x
Big news: Kimi-K3 by @Kimi_Moonshot is now #1 in the Frontend Code Arena with 1679 pts, surpassing Claude Fable 5. This is a 17-place jump from Kimi-k2.6 (#18 -> #1). In Frontend, Kimi-K3 ranked #1 in 6 of 7 domains: Brand & Marketing, Reference-Based Design, Data & Analytics, [i…
-
@arena
@arena
on x
Kimi-K3 just topped the Frontend Code Arena with a 76% pairwise win rate. When its output was compared head-to-head against other models on the same task, it was picked as the better output 76% of the time on average. For reference: Claude Fable 5 (63%), GPT-5.6 Sol (58%). 50% [i…
-
@arena
@arena
on x
In the Text Arena, Kimi-K3 by @Kimi_Moonshot landed #9, with 1486 pts. This is another significant improvement from Kimi-k2.6 (#38 -> #9). - Top 10 in Creative Writing, Coding and Instruction Following - #1 in three occupations: Physical & Social Science, Legal & Government, [ima…
-
@yzhang_cs
Yu Zhang
on x
funny that K3 is great at making 3Blue1Brown videos, and the Quantile Balancing example in the blog took just a few shots to produce. https://kimi.com/...
-
@arena
@arena
on x
Kimi K3 from @Kimi_Moonshot has moved the Pareto Frontier for Code Arena: Frontend. [image]
-
@deredleritt3r
Prinz
on x
The most interesting question about Kimi K3 is whether it poses cyber risk. Kimi K3 benchmarks do not include a CyberGym score. Waiting for @AISecurityInst to bench this model.
-
@yzhang_cs
Yu Zhang
on x
K3 has now crossed the 1M context-length barrier, and DeepSeek's sparse attn has done the same. But what architecture will take us to 5M, 10M, or even longer? I'd always argue that fixed-state linear attn, especially GDN/KDA, is highly competitive here. Hybrid designs are
-
Amin Qurjili
Amin Qurjili
on linkedin
Kimi K3 is out and looks extremely impressive. — But the benchmark numbers may be the least interesting part of the release. …
-
Dave Schroeder, PhD
Dave Schroeder, PhD
on linkedin
This is intentional, by design, and strategically intended in every dimension to elicit exactly this panicked reaction, to destabilize US AI development …
-
Georg Zoeller
Georg Zoeller
on linkedin
The reason this all ends up with war one day is because to Americans everything everyone in the world does is about them and if they are not winning, it's always an attack. …
-
@carnage4life
Dare Obasanjo
on bluesky
More coverage of Kimi K3's achievements below
-
r/neoliberal
r
on reddit
China's open-weight Kimi model stuns AI world with frontier-level results
-
r/StockMarket
r
on reddit
China's open-weight Kimi model stuns AI world with frontier-level results
-
@miles_brundage
Miles Brundage
on x
K3 analysis wen Also, reminder to Americans - we could have this kind of state capacity at home. Let's properly fund and unmuzzle CAISI! https://x.com/...
-
@signulll
@signulll
on x
a while ago there was a device called the palm pre, pretty revolutionary at the time. it ran something entirely new called webOS, which was built using html5. it had multi tasking, card based navigation, was very cool & slick. palm needed hundreds of engineers, years of
-
@aisecurityinst
@aisecurityinst
on x
Our first public analysis of the open/closed weight gap in frontier cyber capabilities finds it is 4-7 months with GLM-5.2 and DeepSeek V4-Pro, narrowing from 6-10 months through most of 2025. Advanced capabilities are reaching less safeguarded open models faster than before. 🧵 […
-
@obrien
Chris O'Brien
on x
In retrospect, was it perhaps a mistake to elect a guy whose primary concern was getting eaten by sharks while on an electric boat to lead the country through this period?
-
@davidsacks
David Sacks
on x
This is concerning. For the first time, a Chinese model Kimi K3 has taken #1 on the Frontend Code Arena and is scoring at or near the frontier on other benchmarks. Meanwhile America is tying itself in knots: politicians and bureaucrats are banning new data centers, piling on
-
@nic_carter
Nic Carter
on x
From an ecological perspective, western frontier models are a common pool resource that Chinese open weight models exploit. They are finite, because if you exploit them too hard, they lose the incentive and resources to train new models, and everyone loses. (Chinese models aren't
-
@mweinbach
Max Weinbach
on x
It finally finished, here's the final output. Used 60% of my monthly Kimi usage on it https://macos27.kimi.page/
-
@pimdewitte
Pim de Witte
on x
kimi is truly a revolutionary model [image]
-
@mweinbach
Max Weinbach
on x
I asked a Kimi K3 Max agent swarm to recreate macOS 27 with real Liquid Glass and native apps in web browser and it's been going for 3 hours [image]
-
@carlquintanilla
Carl Quintanilla
on bluesky
“.. People are worried that if US companies start using Chinese models more and Anthropic less, then Anthropic will invest less. That means those US firms will lower the capex and in the end chip demand will be affected.” — @bloomberg.com — www.bloomberg.com/news/article... …
-
@metacurity.com
Cynthia Brumfield
on bluesky
China's Moonshot just issued its new Kimi K3 model, which the company said rivals the strongest offerings from OpenAI and Anthropic PBC, and this is freaking everyone out, even though it was utterly predictable. — www.bloomberg.com/news/article...
-
r/technology
r
on reddit
China's Moonshot unveils world's largest open AI model, closing in on US rivals
-
@zephyr_z9
@zephyr_z9
on x
“Kimi 3 distilled Fable and therefore big training infra still required for leading edge.” Fable was unbanned on July 1 Generating data, cleaning data, and completing the post-training in 15 days isn't possible
-
@patrickmoorhead
Patrick Moorhead
on x
My initial reaction to Kimi 3: - it's an over-reaction shockingly similar the DeepSeek panic - models like Kimi 3 accelerate and grow the inference market faster than without - Kimi 3 distilled Fable and therefore big training infra still required for leading edge. No Fable, no […
-
@benbajarin
Ben Bajarin
on x
Exactly right. You can argue it's bad for the frontier model labs but all lower token costs does is drastically increase demand for compute. It's the hyperscalers who benefit from this more than anyone and even if THEY were they only ones spending, there will still be more
-
@gavinsbaker
Gavin Baker
on x
Kimi K3 may be an important inflection point for AI. Potentially negative for Anthropic and OpenAI while being net positive for essentially every other company in the world. I mean that very literally. Although the real “Sputnik moment” would be an open-source frontier model that…
-
@quxiaoyin
Xiaoyin Qu
on x
What Kimi K3 means for USA AI: 1. FYI Kimi K3 is open weight and will be released on July 27, 2026. 2. When the best open weight model exceeds the best closed-source model, how does @AnthropicAI justify its Fable pricing? Why would anyone pay for that? Not to mention the Mythos
-
@jukan05
Jukan
on x
Anyone loudly touting how much cheaper Kimi K3 is than Western LLMs is deliberately ignoring how much more expensive it has become compared with previous Chinese LLMs.
-
@jiahanjimliu
Jim Liu
on x
Neoclouds: The Kimi K3 Scare Kimi K3 caused a large scare in the AI trade as this Chinese open source model matched frontier models on benchmarks. Let me unpack what's actually going on. Chinese Labs have much less GPUs than American Labs and yet are able to train “just as
-
@alexfinn
Alex Finn
on x
I was wrong. I said we were a year away from Fable 5 on our desk. That day is today An open model BETTER than Fable 5 in some benchmarks just dropped Better than ChatGPT 5.6 on FrontierSWE. Better than Fable 5 on Automation Bench This fundamentally changes the AI race forever
-
@semianalysis_
@semianalysis_
on x
Kimi K3 2.8T is so large that it will not fit on a single NVIDIA DGX B200, even at FP4. A GB300 NVL72, B300, or MI355X system is required, as each GPU has 288 GB of memory. One optimization that could make Kimi K3 fit on B200 is to gang multiple nodes together and use a [image]
-
@nicky_sap
@nicky_sap
on x
Everyone raving about Kimi K3 so I had to give it a shot. I signed up for free and threw it a pretty nebulous prompt: 'look up Nick Saponaro and create a groundbreaking website using animation, cursor reaction, masterful design, and unique user experience. not just a parallax [vi…
-
@bhavani_00007
@bhavani_00007
on x
I tested Kimi K3 vs Claude Opus 4.8 Same prompt, an armory bay with lighting, props, and detail. Top is Kimi K3, bottom is Opus 4.8. It's not even close. Kimi K3 built a full scene with textures, proper lighting, ammo crates, weapon racks, working detail everywhere. Opus 4.8 [vid…
-
@emollick
Ethan Mollick
on x
Kimi is, as I have been saying, a very good model. But it is not a DeepSeek r1 moment, in that it is roughly where I would expect on the curve rather than an unexpected leap. It will be treated as a DeepSeek moment for a variety of reasons especially as more people hear about it.
-
@emollick
Ethan Mollick
on x
A lot of swift conclusions are being drawn about Kimi K3 based on fairly saturated benchmarks and ELOs, rather than actually testing it on very hard problems. The AI frontier has already moved so far that a good model that is a still months behind looks like the future to many.
-
@deanwball
Dean W. Ball
on x
Some observations on Kimi: 1. It's a very good model! I don't think its performance can be explained away by distillation or anything like that. In agentic coding sessions, it seems pretty much on par with the best public models of Q1 2026. In my fairly limited use, it also
-
@anikasomaia
Anika
on x
a week ago @SemiAnalysis_ wrote that Chinese labs are “simply too compute poor to truly reach the frontier.” today one of those “too compute poor” labs, a 300-person startup actually, shipped a model that compares to opus 4.8 the entire western consensus - export controls, the
-
@shakeelhashim
Shakeel
on x
Kimi K3 threatens to once again send Washington into a panic. But look a little closer, and the hysteria may be unwarranted. On Thursday, the AI industry got its second “DeepSeek moment” — this time courtesy of Moonshot, whose new Kimi K3 model has “erased America's AI lead,”
-
@omarsar0
Elvis
on x
That missing token efficiency is coming. Can't say more now, but efficient frontier long context reasoning/understanding and other TTC breakthroughs are on the horizon. Architectural improvements are often ignored in these benchmarks, but I expect major shifts by EOY.
-
@levie
Aaron Levie
on x
This post is key. The cheaper AI gets, the more opportunity there is for the entire ecosystem - especially including end-customers - to benefit. Everything is bottlenecked by being able to successfully and cost effectively deploy AI in real workloads. Any time we can lower the
-
@chamath
Chamath Palihapitiya
on x
Gavin is right. This is positive for everyone except the closed frontier labs.
-
@samanthaladuc
Samantha LaDuc
on x
A bullish Kimi K3 argument: “Anything that lowers margins and increases competition at the model layer is good for every other AI layer: power, semiconductors, hyperscalers, neoclouds and yes even software.”
-
@liangsays
Brent Liang
on x
one of the best takes on this site re kimi. we are chronicling a mini sputnik moment right now and living through history
-
@pstasiatech
Paul Triolo
on x
...the hypothesis that they have much more advanced model checkpoints internally that are already being used for RSI. In the latter scenario, reaching RSI even a few months ahead of other labs might be enough to cement a permanent lead.
-
@limitingthe
@limitingthe
on x
“Anything that lowers margins and increases competition at the model layer is good for every other AI layer: power, semiconductors, hyperscalers, neoclouds and yes even software.”
-
@notthreadguy
@notthreadguy
on x
genuinely one of the best bull posts I've ever read > China open source ai is better than US? Long all capex beneficiaries > China open source ai is NOT better than US? Long all capex beneficiaries
-
@shanumathew93
Shanu Mathew
on x
Great insights. Gavin nails it. If an oligopoloy of labs sustain 90% inference margins, they capture most of the economics and eventually vertically integrate the stack & squeeze chips, power, data centers, cloud and software. More competition at the model layer —> lower model
-
@bare_birk
Birk
on x
Interesting post from @GavinSBaker! Kimi K3 is bad for Anthropic and OpenAI but good for all other companies. Margins will go from the frontier labs, to all other companies in the sector. Infrastructure will still be very important in both scenarios (Opensource vs closed source)
-
@dzambhalahodl
Steven Lubka
on x
Kimi3 is a good model, but it is not a “cheap” model nor is it a “small” model. Kimi3 showed that China can train a decent competing model less than 6 months behind the US. It did not show that they can compete with Frontier at a similar efficiency advantage to Deepseek etc.
-
r/technology
r
on reddit
Chinese startup Moonshot AI unveils Kimi model it says rivals OpenAI, Anthropic