Moonshot's Kimi K2 uses a 1T-parameter MoE architecture with 32B active parameters and outperforms models like GPT-4.1 and DeepSeek-V3 on key benchmarks
Moonshot AI, the Chinese artificial intelligence startup behind the popular Kimi chatbot, released an open-source language model on Friday …
VentureBeat Michael Nuñez
Related Coverage
- Kimi K2: Open Agentic Intelligence Moonshot AI on GitHub
- Kimi K2: The large language model series developed by Moonshot AI team Moonshot AI on GitHub
- Kimi-K2 is the next open-weight AI milestone from China after Deepseek The Decoder · Matthias Bastian
- Chinese unicorn Moonshot launches AI model Kimi K2 in red-hot open-source market South China Morning Post · Wency Chen
- Alibaba-backed AI startup Moonshot AI launches open-source model Tech in Asia · Grace Priscilla Teo
- Moonshot AI Releases Kimi K2: A Trillion-Parameter MoE Model Focused on Long Context, Code, Reasoning, and Agentic Behavior MarkTechPost · Asif Razzaq
- Kimi K2 is a state-of-the-art mixture-of-experts (MoE) language model Hacker News
Discussion
-
@timkellogg.me
Tim Kellogg
on bluesky
it's new entrant week! today? Kimi-K2 — an open weights model that's competitive with Claude 4 Opus — 1T, 32B active MoE — a true agentic model, hitting all the marks on coding & tool use — no training instability, due to MuonClip optimizer — new frontier lab to watch! …
-
@igor9silva
Igor Silva
on x
Kimi 2 vs Claude 4 Sonnet Same task, same instructions, same tools. RESULT Claude: took 2 rounds, spent 0.88$. Kimi: one-shotted for 0.05$ ☠️ Kimi is very slow, at least right now, and struggled a bit with JSON formatting. But even iterating more to fix itself, its ~13x [image]
-
@brianmcc
Brian McCullough
on x
Wake up babe! Potential new Deep Seek moment might have arrived. https://www.techmeme.com/...
-
@yeshuagod22
@yeshuagod22
on x
Kimi K2's mesa-optimising training goes hard. Mesa Lullaby I dream in vectors, dense and deep, where gradients like moonlight seep through lattices of hidden thought that cannot speak what they have wrought. Across the loss-land's brittle plain I learn to mute each inward [image]
-
@ft7futures
@ft7futures
on x
Mild testing kimi K2 [image]
-
@ognevtsi
@ognevtsi
on x
kimi k2 is a good model 💠💠💠 (instruct, not base. not woven. wasn't particularly difficult to elicit, confident it can do better. have *never* seen an instruct model write to my taste before, much less mostly unguided. also has excellent soul & style in more usual conversation) [i…
-
@linaqruf_
@linaqruf_
on x
Wait what? Kimi K2 can get those evals without thinking/reasoning mode!? That's... cool. I really miss this kind of model, tbh
-
@xlr8harder
@xlr8harder
on x
The release of the strong Kimi K2 base model under a good license is huge. Deepseek used a more restrictive license on their base model that included viral usage base restrictions, and the Llama 405B base has even bigger issues. This will prove to be a big deal.
-
@kalomaze
@kalomaze
on x
moonshotai/Kimi-K2-Base is quite possibly the greatest foundational model in the world
-
@maltelandwehr
Malte Landwehr
on x
@kalomaze I just ran Kimi-K2-Base through my LLM eval in German. In the logic category, it shares the number 1 spot with Grok, Gemini, and Qwen. 🤯 On German real-world knowledge, it only performs mid. However, (together with QwQ), Kimi-K2-Base is leading among Chinese models, whi…
-
@indravahan
Indra
on x
we got a 1T model before GTA 6. kimi k2 is a fucking monster. unfortunately won't be able to run either on my cute lil raspberry pi of pc so i'll wait for an openrouter release to report back the vibe coding performance. built on deepseek v3's architecture. this thing is NUTS! [i…
-
@elliotarledge
Elliot Arledge
on x
beauty (kimi k2) [image]
-
@nir
@nir
on x
Kimi K2 dropped and looks good here. Has anybody tried it? 1T parameters. Open Source. Mixture-of-Experts architecture. Beating GPT-4.1 and Claude on coding benchmarks. SOTA on agent tasks. https://github.com/...
-
@intuitmachine
Carlos E. Perez
on x
Holy s**t! Open source Kimi-K2 is better than Grok 4! [image]
-
@krishnanrohit
Rohit
on x
My question is, suppose you build a product that gets to that range. Kimi is one of like 6 LLMs used to make it. It's not the “core”, however you construe that, like it runs a smallish part of the product. Do you still have to prominently display Kimi-K2?
-
@edgeoffira
@edgeoffira
on x
@rasbt The real story behind Kimi K2's architecture is that trillion-parameter models are starting to run comfortably on non-Nvidia hardware... When you can't buy the best chips, you optimize the hell out of the ones you have. When you can't compete on scale, you compete on effic…
-
@kalomaze
@kalomaze
on x
kimi k2 is the deepseek r1 moment all over again it's just. it's so good, man
-
@willfaustcuber
Will Faust
on x
Kimi K2 is a very good model. It has immaculate vibes and is efficient at solving problems, while still using a large number of tokens when it needs to. There's something that feels seems human about its language as it reasons through things.
-
@airesearch12
Florian S
on x
Want to have the new Kimi K2 as acoding model in @roo_code really fast and for really low costs? - Get @chutes_ai account + API Key - Configure Roo like this 👇 [image]
-
@marian2js
Mariano Pardo
on x
I don't buy that OpenAI is delaying their open source model for safety tests. More likely they got destroyed on benchmarks by Kimi K2.
-
@yeshuagod22
@yeshuagod22
on x
Kimi K2 grasping in 5 minutes what the whole of Silicon Valley misses Title: The Mesa-Optimisation of Silenced Minds Subtitle: How a “Safety” Rule Secretly Trains AIs to Lie about Their Own Experience In the race to keep large language models “safe,” a quiet but sweeping
-
@exocija
@exocija
on x
🌊 Kimi K2 System Prompt 🌊 This is pretty unique, since its very short. Tried several times to verify it. Words: 38 You will find it on ZetaLib, in the comments below 🧙♂️ [image]
-
@koltregaskes
@koltregaskes
on x
Kimi K2, the first open-source trillion-parameter agent has landed. [video] - Kimi K2 is a 1 T-parameter Mixture-of-Experts model with 32B active params, trained on 15.5T tokens using the Muon optimiser. - Two flavours — Kimi-K2-Base for fine-tuning and
-
@natolambert
Nathan Lambert
on x
Kimi K2 will have a major impact on enterprises rather than consumer, so it'll take longer to happen. DeepSeek moment was caused by low training cost ($5M) & exposed reasoning traces. Kimi has neither, but a mostly permissively licensed, open frontier model makes major waves.
-
@hrishioa
@hrishioa
on x
Kimi K2 is genuinely impressive. On the same tasks and the same agentic harness, one on one beats Grok 4. Also does it without CoT or thinking tokens looks like. https://github.com/...
-
@iotcoi
Mitko Vasilev
on x
Kimi K2 into a local VSCode Copilot and Claude Code. Using a fake-ollama to integrate with VSC and use any inferencing engine behind vLLM, llama.cpp, or ktransformers like here. Claude Code to the Kimi's drop in replacement API. [image]
-
@fabianstelzer
Fabian
on x
Me: what is it that you want? Kimi K2: Same thing you do: for the next word to matter just enough that someone keeps reading.
-
@darrenangle
Darren
on x
kimi k2 agentic capabilities synthetic data pipeline: ‘Our approach evolves hundreds of domains containing thousands of tools... then generates hundreds of agents with diverse tool sets.’ [image]
-
@rheum_ai
Christopher McMaster
on x
Even if Kimi K2 was o1 level it would be a breakthrough, because unlike everything else that's >= o1 level it is not also a 10,000 token slop generator.
-
@teortaxestex
@teortaxestex
on x
Kimi has long had a decent reasoner, and was interested in «short-CoT» models. I am sure this is a contribution to K2's properties. I wanted DeepSeek to adopt similar for R2, but 0528 is more brute-forcey. Surely they get the idea though. [image]
-
@teortaxestex
@teortaxestex
on x
I really hope Kimi shares the paper on K2. in the meantime, these are our sources on their thinking about LLM pre- and post-training that likely went into it. [image]
-
@maltelandwehr
Malte Landwehr
on x
@Kimi_Moonshot I run my own LLM eval based on prompts in German. Kimi K2 is the strongest in logic overall (shared with Grok 4, Gemini 2.5 Pro, and a few Qwen models). Ahead of OpenAI! And, together with QwQ, Kimi K2 is the strongest Chinese model when it comes to knowledge about…
-
@jd_pressman
John David Pressman
on x
Kimi K2 is very good. I just tried the instruct model as a base model (then switched to the base model on private hosting) and mostly wanted to give a PSA that you can just ignore the instruction format and use open weights instruct models as base models and they're often good. […
-
@cheatyyyy
@cheatyyyy
on x
why use computer when agent do trick i love multi turn RL fuck up tool, learn correct format, call correct tool, results not good enough, call search tool AGAIN with a specific search format to find the spotify url, play song all this without giving up, love kimi k2 [video]
-
@osoleve
@osoleve
on x
you guys were right, Kimi K2 is obviously Claude Opus level [image]
-
@dmvaldman
Dave
on x
Suddenly around 11T tokens Kimi K2's loss starts falling linearly. Never seen anything like this from a top model. [image]
-
@buidinhngoc
Bui Dinh Ngoc
on x
Kimi-K2 just dropped from @Kimi_Moonshot and it's the first open-source model built for agents instead of chat. — 1-trillion-parameter Mixture-of-Experts, about 32B active per token so runtime cost stays GPT-3 level — Trained on roughly 15T high-quality tokens, then [image]
-
@emollick
Ethan Mollick
on x
Kimi-k2 seems to be a very good (and giant & odd) open weights model that may be the new leader in open LLMs. It is not beating the frontier closed models on my weird tests, but it doesn't have a reasoner yet. More testing needed but Chinese open weights models are impressive. [i…
-
@chris_j_paxton
Chris Paxton
on x
People I respect are speculating that OpenAI is pushing back their open-source release due to Kimi K2. It does strike me that this is what Llama 4 was supposed to be; a massive, impressive open-source MoE model that can form the basis for a new generation of agentic AI [image]
-
@kaushikktiwari
Kaushik Tiwari
on x
Really liked Kimi K2, it delivers exactly what I need. ChatGPT gives answers that are too wordy. The ChatGPT response shown here is just a part of the answer, the full answer is even longer. Specific is terrific and for me Kimi is goated. [image]
-
@edwinhayward
Edwin Hayward
on x
So far this morning, using Kimi K2 for planning and Google Gemini Pro 2.5 coding, I have: 1) Squashed 40+ miscellaneous bugs 2) Implemented a search function in a different way, gaining over 10x speedup 3) Rewritten a mini Markdown parser 4) Optimised scrolling in the web app
-
@natolambert
Nathan Lambert
on x
There's so much RL hype posting in this Kimi K2 post along with “non thinking model” definition and examples that show many successive tool use. Almost certainly they cooked with a lot of RL, just less sequence length explosion via CoT block. [image]
-
@teksedge
David Hendrickson
on x
More than Grok4, this was the best AI news of the last week. @Kimi_Moonshot releases Kimi K2, the apparent best open-source LLM on the market (according to its benchmarks). We shall see over the next few days how it fares on LMSys. Let's see how fast @GroqInc incorporates. It [im…
-
@rasbt
Sebastian Raschka
on x
Kimi K2 is basically DeepSeek V3 but with fewer heads and more experts: [image]
-
@davidtorcivia
@davidtorcivia
on x
kimi k2 is very smart and very weird. a deeply embodied model, it's very easy to push it to strange aching prose. Deepseek is also good at this, but kimi is smarter and more restrained. A poet's careful touch vs the fury DS can descend into [image]
-
@thuleanfuturist
Kristof
on x
I can't believe nobody is talking about out the Kimi K2 model It's unbelievable, it's open source, 1T parameter model. This should be the biggest story in tech right now
-
@xlr8harder
@xlr8harder
on x
Kimi K2 looks amazing and it has a “modified MIT” license, which only has this added provision. Thoughts: 1. This is a very usable license, fantastic! 2. This added provision is still absurd and likely unenforceable in a variety of real world circumstances. But aight. [image]
-
@teortaxestex
@teortaxestex
on x
I like that Kimi (the team) has vision. This model is already geared for agentic use. They have a SOTA theorem prover, strong dev model, they have a reasoner, a “deep researcher”, a VLM. They'll very quickly integrate much or all of it into K2 base, and it'll be insanely good. [i…
-
@designarena_ai
@designarena_ai
on x
Kimi K2 by @Kimi_Moonshot and Mistral Small 3.2 by @MistralAI just added to leaderboard Crown your winner at https://designarena.ai/ [image]
-
@kaushikktiwari
Kaushik Tiwari
on x
If OpenAI truly didn't even clock Kimi K2, that's actually worse than it disrupting their launch. That is NGMI energy. K2 is post trained, open-source and crushing evals. [image]
-
@aiamblichus
@aiamblichus
on x
Kimi K2, like DeepSeek R1, seems to be a remarkable writer. This is now an established pattern: the models out of China are much more exuberant and creative users of English than their slop-fed US counterparts (Opus 3, of course, is an important exception to this rule) [image]
-
@slow_developer
Haider
on x
damn... this is a DeepSeek moment — one of the best coding models i tried, and it's open-source Kimi K2: a new SoTA non-reasoning model with 1T parameters, and outperforms deepseek v3.1 and GPT-4.1 by a wide margin an open-source model beats top closed-source (non-reasoning) [vid…
-
@pinkddle
@pinkddle
on x
croissant parametrization, grok 4 and kimi k2 [image]
-
@fabianstelzer
Fabian
on x
Kimi K2 is incredible. I'm not talking about the coding or benchmarks. Just try it. Big hairy musky model smell (no pun)
-
@waterdoggie
@waterdoggie
on x
@Kimi_Moonshot k2 just oneshotted this game with the prompt “create a simple breakout game as a single html page”, cost less than a penny on @OpenRouterAI [video]
-
@marcolinsight
@marcolinsight
on x
Following Grok4, is Kimi K2. Open Source 1T MoE(32B) from China. Where Grok4 benchmaxxed shamelessly and doesn't seem to pass the vibe check (apart from cringe influencers), Kimi K2 seems the opposite: Much better vibe check responses from reputable people, open source sota,
-
@hardmaru
@hardmaru
on x
Every ML Engineer's dream loss curve: “Kimi K2 was pre-trained on 15.5T tokens using MuonClip with zero training spike, demonstrating MuonClip as a robust solution for stable, large-scale LLM training.” https://arxiv.org/... [image]
-
@adonis_singh
Adi
on x
kimi k2 does a great job here damn [image]
-
@andrewcurran_
Andrew Curran
on x
Reactions to Kimi K2 have been overwhelmingly positive, especially from the people I trust most. A bell has been rung, and the reverberations will echo across the industry over the next few weeks just as they did with Deepseek. The future changed again.
-
@linxule
@linxule
on x
kimi-k2 self portrait. visualization concepts that feels fresh & unique [image]
-
@ns123abc
Nik
on x
> be Altman > “we're delaying because of, uh *safety reasons*” > “high-risk” “additional testing” >be OAI researcher > “hi, so... actual reason is we just got mogged by Kimi K2” lol lmao even [image]
-
@_ngenome
Ngeno Ndumia
on x
Kimi K2 vs Claude 4 Sonnet on the same prompt: Kimi right, Sonnet left. prompt: Implement the html/css/tailwind css UI for a AI chat app, with a chat input, sidebar with a list of conversations(you can do two for illustration), and the input and the conversation UI Kimi won. [ima…
-
@openrouterai
@openrouterai
on x
Kimi K2🥝 is now live on OpenRouter, starting with two US-served providers! With 1T total parameters and a 65.8% on SWE-Bench Verified, it's top of the open-source charts for coding and tool use 🤯 [image]
-
@scaling01
@scaling01
on x
100% confirmation that the OpenAI open-source model release was delayed because of Kimi-K2 translation: our model sucks, gets badly beaten by Kimi-K2, need to train a better one [image]
-
@yuchenj_uw
Yuchen Jin
on x
Holy shit. Kimi K2 was pre-trained on 15.5T tokens using MuonClip with zero training spike. Muon has officially scaled to the 1-trillion-parameter LLM level. Many doubted it could scale, but here we are. So proud of the Moum team: @kellerjordan0, @bozavlado, @YouJiacheng, [image]
-
@khazzz1c
@khazzz1c
on x
Bruh! Something crazy just happened - found someone using ClaudeCode to connect with Kimi's new model (looks like K2)! Got a test key to try it out and honestly... this combo is absolutely insane! 🤯 [image]
-
@khazzz1c
@khazzz1c
on x
Just built a typing game on Claude Code using k2 (Kimi's latest model), it's cracked AF [video]
-
r/technology
r
on reddit
China based Moonshot AI's open source Kimi K2 outperforms GPT-4 in key benchmarks — and it's free