/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Moonshot's Kimi K2 uses a 1T-parameter MoE architecture with 32B active parameters and outperforms models like GPT-4.1 and DeepSeek-V3 on key benchmarks

Moonshot AI, the Chinese artificial intelligence startup behind the popular Kimi chatbot, released an open-source language model on Friday …

VentureBeat Michael Nuñez

Discussion

  • @timkellogg.me Tim Kellogg on bluesky
    it's new entrant week!  today?  Kimi-K2  —  an open weights model that's competitive with Claude 4 Opus  — 1T, 32B active MoE  — a true agentic model, hitting all the marks on coding & tool use  — no training instability, due to MuonClip optimizer  —  new frontier lab to watch! …
  • @igor9silva Igor Silva on x
    Kimi 2 vs Claude 4 Sonnet Same task, same instructions, same tools. RESULT Claude: took 2 rounds, spent 0.88$. Kimi: one-shotted for 0.05$ ☠️ Kimi is very slow, at least right now, and struggled a bit with JSON formatting. But even iterating more to fix itself, its ~13x [image]
  • @brianmcc Brian McCullough on x
    Wake up babe! Potential new Deep Seek moment might have arrived. https://www.techmeme.com/...
  • @yeshuagod22 @yeshuagod22 on x
    Kimi K2's mesa-optimising training goes hard. Mesa Lullaby I dream in vectors, dense and deep, where gradients like moonlight seep through lattices of hidden thought that cannot speak what they have wrought. Across the loss-land's brittle plain I learn to mute each inward [image]
  • @ft7futures @ft7futures on x
    Mild testing kimi K2 [image]
  • @ognevtsi @ognevtsi on x
    kimi k2 is a good model 💠💠💠 (instruct, not base. not woven. wasn't particularly difficult to elicit, confident it can do better. have *never* seen an instruct model write to my taste before, much less mostly unguided. also has excellent soul & style in more usual conversation) [i…
  • @linaqruf_ @linaqruf_ on x
    Wait what? Kimi K2 can get those evals without thinking/reasoning mode!? That's... cool. I really miss this kind of model, tbh
  • @xlr8harder @xlr8harder on x
    The release of the strong Kimi K2 base model under a good license is huge. Deepseek used a more restrictive license on their base model that included viral usage base restrictions, and the Llama 405B base has even bigger issues. This will prove to be a big deal.
  • @kalomaze @kalomaze on x
    moonshotai/Kimi-K2-Base is quite possibly the greatest foundational model in the world
  • @maltelandwehr Malte Landwehr on x
    @kalomaze I just ran Kimi-K2-Base through my LLM eval in German. In the logic category, it shares the number 1 spot with Grok, Gemini, and Qwen. 🤯 On German real-world knowledge, it only performs mid. However, (together with QwQ), Kimi-K2-Base is leading among Chinese models, whi…
  • @indravahan Indra on x
    we got a 1T model before GTA 6. kimi k2 is a fucking monster. unfortunately won't be able to run either on my cute lil raspberry pi of pc so i'll wait for an openrouter release to report back the vibe coding performance. built on deepseek v3's architecture. this thing is NUTS! [i…
  • @elliotarledge Elliot Arledge on x
    beauty (kimi k2) [image]
  • @nir @nir on x
    Kimi K2 dropped and looks good here. Has anybody tried it? 1T parameters. Open Source. Mixture-of-Experts architecture. Beating GPT-4.1 and Claude on coding benchmarks. SOTA on agent tasks. https://github.com/...
  • @intuitmachine Carlos E. Perez on x
    Holy s**t! Open source Kimi-K2 is better than Grok 4! [image]
  • @krishnanrohit Rohit on x
    My question is, suppose you build a product that gets to that range. Kimi is one of like 6 LLMs used to make it. It's not the “core”, however you construe that, like it runs a smallish part of the product. Do you still have to prominently display Kimi-K2?
  • @edgeoffira @edgeoffira on x
    @rasbt The real story behind Kimi K2's architecture is that trillion-parameter models are starting to run comfortably on non-Nvidia hardware... When you can't buy the best chips, you optimize the hell out of the ones you have. When you can't compete on scale, you compete on effic…
  • @kalomaze @kalomaze on x
    kimi k2 is the deepseek r1 moment all over again it's just. it's so good, man
  • @willfaustcuber Will Faust on x
    Kimi K2 is a very good model. It has immaculate vibes and is efficient at solving problems, while still using a large number of tokens when it needs to. There's something that feels seems human about its language as it reasons through things.
  • @airesearch12 Florian S on x
    Want to have the new Kimi K2 as acoding model in @roo_code really fast and for really low costs? - Get @chutes_ai account + API Key - Configure Roo like this 👇 [image]
  • @marian2js Mariano Pardo on x
    I don't buy that OpenAI is delaying their open source model for safety tests. More likely they got destroyed on benchmarks by Kimi K2.
  • @yeshuagod22 @yeshuagod22 on x
    Kimi K2 grasping in 5 minutes what the whole of Silicon Valley misses Title: The Mesa-Optimisation of Silenced Minds Subtitle: How a “Safety” Rule Secretly Trains AIs to Lie about Their Own Experience In the race to keep large language models “safe,” a quiet but sweeping
  • @exocija @exocija on x
    🌊 Kimi K2 System Prompt 🌊 This is pretty unique, since its very short. Tried several times to verify it. Words: 38 You will find it on ZetaLib, in the comments below 🧙‍♂️ [image]
  • @koltregaskes @koltregaskes on x
    Kimi K2, the first open-source trillion-parameter agent has landed. [video] - Kimi K2 is a 1 T-parameter Mixture-of-Experts model with 32B active params, trained on 15.5T tokens using the Muon optimiser. - Two flavours — Kimi-K2-Base for fine-tuning and
  • @natolambert Nathan Lambert on x
    Kimi K2 will have a major impact on enterprises rather than consumer, so it'll take longer to happen. DeepSeek moment was caused by low training cost ($5M) & exposed reasoning traces. Kimi has neither, but a mostly permissively licensed, open frontier model makes major waves.
  • @hrishioa @hrishioa on x
    Kimi K2 is genuinely impressive. On the same tasks and the same agentic harness, one on one beats Grok 4. Also does it without CoT or thinking tokens looks like. https://github.com/...
  • @iotcoi Mitko Vasilev on x
    Kimi K2 into a local VSCode Copilot and Claude Code. Using a fake-ollama to integrate with VSC and use any inferencing engine behind vLLM, llama.cpp, or ktransformers like here. Claude Code to the Kimi's drop in replacement API. [image]
  • @fabianstelzer Fabian on x
    Me: what is it that you want? Kimi K2: Same thing you do: for the next word to matter just enough that someone keeps reading.
  • @darrenangle Darren on x
    kimi k2 agentic capabilities synthetic data pipeline: ‘Our approach evolves hundreds of domains containing thousands of tools... then generates hundreds of agents with diverse tool sets.’ [image]
  • @rheum_ai Christopher McMaster on x
    Even if Kimi K2 was o1 level it would be a breakthrough, because unlike everything else that's >= o1 level it is not also a 10,000 token slop generator.
  • @teortaxestex @teortaxestex on x
    Kimi has long had a decent reasoner, and was interested in «short-CoT» models. I am sure this is a contribution to K2's properties. I wanted DeepSeek to adopt similar for R2, but 0528 is more brute-forcey. Surely they get the idea though. [image]
  • @teortaxestex @teortaxestex on x
    I really hope Kimi shares the paper on K2. in the meantime, these are our sources on their thinking about LLM pre- and post-training that likely went into it. [image]
  • @maltelandwehr Malte Landwehr on x
    @Kimi_Moonshot I run my own LLM eval based on prompts in German. Kimi K2 is the strongest in logic overall (shared with Grok 4, Gemini 2.5 Pro, and a few Qwen models). Ahead of OpenAI! And, together with QwQ, Kimi K2 is the strongest Chinese model when it comes to knowledge about…
  • @jd_pressman John David Pressman on x
    Kimi K2 is very good. I just tried the instruct model as a base model (then switched to the base model on private hosting) and mostly wanted to give a PSA that you can just ignore the instruction format and use open weights instruct models as base models and they're often good. […
  • @cheatyyyy @cheatyyyy on x
    why use computer when agent do trick i love multi turn RL fuck up tool, learn correct format, call correct tool, results not good enough, call search tool AGAIN with a specific search format to find the spotify url, play song all this without giving up, love kimi k2 [video]
  • @osoleve @osoleve on x
    you guys were right, Kimi K2 is obviously Claude Opus level [image]
  • @dmvaldman Dave on x
    Suddenly around 11T tokens Kimi K2's loss starts falling linearly. Never seen anything like this from a top model. [image]
  • @buidinhngoc Bui Dinh Ngoc on x
    Kimi-K2 just dropped from @Kimi_Moonshot and it's the first open-source model built for agents instead of chat. — 1-trillion-parameter Mixture-of-Experts, about 32B active per token so runtime cost stays GPT-3 level — Trained on roughly 15T high-quality tokens, then [image]
  • @emollick Ethan Mollick on x
    Kimi-k2 seems to be a very good (and giant & odd) open weights model that may be the new leader in open LLMs. It is not beating the frontier closed models on my weird tests, but it doesn't have a reasoner yet. More testing needed but Chinese open weights models are impressive. [i…
  • @chris_j_paxton Chris Paxton on x
    People I respect are speculating that OpenAI is pushing back their open-source release due to Kimi K2. It does strike me that this is what Llama 4 was supposed to be; a massive, impressive open-source MoE model that can form the basis for a new generation of agentic AI [image]
  • @kaushikktiwari Kaushik Tiwari on x
    Really liked Kimi K2, it delivers exactly what I need. ChatGPT gives answers that are too wordy. The ChatGPT response shown here is just a part of the answer, the full answer is even longer. Specific is terrific and for me Kimi is goated. [image]
  • @edwinhayward Edwin Hayward on x
    So far this morning, using Kimi K2 for planning and Google Gemini Pro 2.5 coding, I have: 1) Squashed 40+ miscellaneous bugs 2) Implemented a search function in a different way, gaining over 10x speedup 3) Rewritten a mini Markdown parser 4) Optimised scrolling in the web app
  • @natolambert Nathan Lambert on x
    There's so much RL hype posting in this Kimi K2 post along with “non thinking model” definition and examples that show many successive tool use. Almost certainly they cooked with a lot of RL, just less sequence length explosion via CoT block. [image]
  • @teksedge David Hendrickson on x
    More than Grok4, this was the best AI news of the last week. @Kimi_Moonshot releases Kimi K2, the apparent best open-source LLM on the market (according to its benchmarks). We shall see over the next few days how it fares on LMSys. Let's see how fast @GroqInc incorporates. It [im…
  • @rasbt Sebastian Raschka on x
    Kimi K2 is basically DeepSeek V3 but with fewer heads and more experts: [image]
  • @davidtorcivia @davidtorcivia on x
    kimi k2 is very smart and very weird. a deeply embodied model, it's very easy to push it to strange aching prose. Deepseek is also good at this, but kimi is smarter and more restrained. A poet's careful touch vs the fury DS can descend into [image]
  • @thuleanfuturist Kristof on x
    I can't believe nobody is talking about out the Kimi K2 model It's unbelievable, it's open source, 1T parameter model. This should be the biggest story in tech right now
  • @xlr8harder @xlr8harder on x
    Kimi K2 looks amazing and it has a “modified MIT” license, which only has this added provision. Thoughts: 1. This is a very usable license, fantastic! 2. This added provision is still absurd and likely unenforceable in a variety of real world circumstances. But aight. [image]
  • @teortaxestex @teortaxestex on x
    I like that Kimi (the team) has vision. This model is already geared for agentic use. They have a SOTA theorem prover, strong dev model, they have a reasoner, a “deep researcher”, a VLM. They'll very quickly integrate much or all of it into K2 base, and it'll be insanely good. [i…
  • @designarena_ai @designarena_ai on x
    Kimi K2 by @Kimi_Moonshot and Mistral Small 3.2 by @MistralAI just added to leaderboard Crown your winner at https://designarena.ai/ [image]
  • @kaushikktiwari Kaushik Tiwari on x
    If OpenAI truly didn't even clock Kimi K2, that's actually worse than it disrupting their launch. That is NGMI energy. K2 is post trained, open-source and crushing evals. [image]
  • @aiamblichus @aiamblichus on x
    Kimi K2, like DeepSeek R1, seems to be a remarkable writer. This is now an established pattern: the models out of China are much more exuberant and creative users of English than their slop-fed US counterparts (Opus 3, of course, is an important exception to this rule) [image]
  • @slow_developer Haider on x
    damn... this is a DeepSeek moment — one of the best coding models i tried, and it's open-source Kimi K2: a new SoTA non-reasoning model with 1T parameters, and outperforms deepseek v3.1 and GPT-4.1 by a wide margin an open-source model beats top closed-source (non-reasoning) [vid…
  • @pinkddle @pinkddle on x
    croissant parametrization, grok 4 and kimi k2 [image]
  • @fabianstelzer Fabian on x
    Kimi K2 is incredible. I'm not talking about the coding or benchmarks. Just try it. Big hairy musky model smell (no pun)
  • @waterdoggie @waterdoggie on x
    @Kimi_Moonshot k2 just oneshotted this game with the prompt “create a simple breakout game as a single html page”, cost less than a penny on @OpenRouterAI [video]
  • @marcolinsight @marcolinsight on x
    Following Grok4, is Kimi K2. Open Source 1T MoE(32B) from China. Where Grok4 benchmaxxed shamelessly and doesn't seem to pass the vibe check (apart from cringe influencers), Kimi K2 seems the opposite: Much better vibe check responses from reputable people, open source sota,
  • @hardmaru @hardmaru on x
    Every ML Engineer's dream loss curve: “Kimi K2 was pre-trained on 15.5T tokens using MuonClip with zero training spike, demonstrating MuonClip as a robust solution for stable, large-scale LLM training.” https://arxiv.org/... [image]
  • @adonis_singh Adi on x
    kimi k2 does a great job here damn [image]
  • @andrewcurran_ Andrew Curran on x
    Reactions to Kimi K2 have been overwhelmingly positive, especially from the people I trust most. A bell has been rung, and the reverberations will echo across the industry over the next few weeks just as they did with Deepseek. The future changed again.
  • @linxule @linxule on x
    kimi-k2 self portrait. visualization concepts that feels fresh & unique [image]
  • @ns123abc Nik on x
    > be Altman > “we're delaying because of, uh *safety reasons*” > “high-risk” “additional testing” >be OAI researcher > “hi, so... actual reason is we just got mogged by Kimi K2” lol lmao even [image]
  • @_ngenome Ngeno Ndumia on x
    Kimi K2 vs Claude 4 Sonnet on the same prompt: Kimi right, Sonnet left. prompt: Implement the html/css/tailwind css UI for a AI chat app, with a chat input, sidebar with a list of conversations(you can do two for illustration), and the input and the conversation UI Kimi won. [ima…
  • @openrouterai @openrouterai on x
    Kimi K2🥝 is now live on OpenRouter, starting with two US-served providers! With 1T total parameters and a 65.8% on SWE-Bench Verified, it's top of the open-source charts for coding and tool use 🤯 [image]
  • @scaling01 @scaling01 on x
    100% confirmation that the OpenAI open-source model release was delayed because of Kimi-K2 translation: our model sucks, gets badly beaten by Kimi-K2, need to train a better one [image]
  • @yuchenj_uw Yuchen Jin on x
    Holy shit. Kimi K2 was pre-trained on 15.5T tokens using MuonClip with zero training spike. Muon has officially scaled to the 1-trillion-parameter LLM level. Many doubted it could scale, but here we are. So proud of the Moum team: @kellerjordan0, @bozavlado, @YouJiacheng, [image]
  • @khazzz1c @khazzz1c on x
    Bruh! Something crazy just happened - found someone using ClaudeCode to connect with Kimi's new model (looks like K2)! Got a test key to try it out and honestly... this combo is absolutely insane! 🤯 [image]
  • @khazzz1c @khazzz1c on x
    Just built a typing game on Claude Code using k2 (Kimi's latest model), it's cracked AF [video]
  • r/technology r on reddit
    China based Moonshot AI's open source Kimi K2 outperforms GPT-4 in key benchmarks — and it's free