/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Cursor releases Composer 2.5, saying it's better at sustained work on long-running tasks and follows complex instructions more reliably; it's built on Kimi K2.5

Composer 2.5 is now available in Cursor.  —  It's a substantial improvement in intelligence and behavior over Composer 2.

Cursor

Context & Ripple Effects

Cursor’s Composer line moved from the March launch of Composer 2—positioned for autonomous, lengthy coding work—to an updated release shortly afterward. Moonshot and Cursor had separately clarified that Composer 2 began from Kimi K2.5, with access provided through Fireworks AI.

The update arrives as Cursor faces scrutiny over AI-assisted coding reliability, while its CEO has cautioned that “vibe coding” can leave advanced projects on unstable foundations. That makes sustained-task execution and instruction-following central product claims rather than merely incremental model benchmarks.

First-order effects

  • Cursor users gain Composer 2.5 in the product, with Cursor asserting better performance on long-running coding tasks and complex instructions.
  • Cursor further ties the Composer offering to Kimi K2.5, reinforcing Moonshot’s model as an underlying component of Cursor’s coding-agent stack.

Second-order effects

  • Cursor will be judged more directly on whether its coding agents can maintain correctness across extended workflows, particularly given the existing reliability criticism around AI-assisted development.
  • The release raises the competitive bar for coding-agent providers: product differentiation shifts toward dependable multi-step execution and adherence to detailed requirements, not simply code-generation quality.

Third-order effects

  • If these improvements prove durable, AI coding tools may increasingly compete as autonomous workflow systems, making reliability, supervision, and maintainability more consequential than isolated model capability claims.
  • Cursor’s dependence on an external foundation model illustrates a layered market in which end-user agent companies differentiate through training, behavior, and product integration while relying on upstream model providers and inference partners.

The trend: This is part of the shift from AI coding assistants that generate snippets toward agents expected to carry out longer, instruction-heavy software tasks with dependable behavior.

Discussion

  • @cursor_ai @cursor_ai on x
    Introducing Composer 2.5, our most powerful model yet. It's more intelligent, better at sustained work on long-running tasks, and more reliable at following complex instructions. For the next week, we're doubling the included usage of the model. [image]
  • @elonmusk Elon Musk on x
    Try it out! (Partially trained on Colossus 2)
  • @mntruell Michael Truell on x
    Composer 2.5 is a significant step up from Composer 2. This is the very start of our work with SpaceXAI. Hope to have more improvements out soon.
  • @sjwhitmore Sam Whitmore on x
    composer 2.5 is really really great. I had it on last week for some testing, forgot that it was on, & totally didn't realize I wasn't on gpt 5.5 (my usual) for a while. the team did a fantastic job!!
  • @cursor_ai @cursor_ai on x
    Together with SpaceXAI, we're training a significantly larger model from scratch, using 10x more total compute. With Colossus 2's million H100-equivalents and our combined data and training techniques, we expect this to be a major leap in model capability.
  • @cursor_ai @cursor_ai on x
    We improved Composer by scaling training, generating more complex RL environments, and introducing new learning methods. For example, we use text feedback during RL to learn faster by assigning credit in rollouts spanning hundreds of thousands of tokens.
  • @cursor_ai @cursor_ai on x
    Composer 2.5 is built on the same open-source base as Composer 2, Moonshot's Kimi K2.5. [image]
  • @cursor_ai @cursor_ai on x
    Composer 2.5 is exceptionally intelligent and up to 10x more efficient than similarly capable models. [image]
  • @mil000 Milo Smith on x
    Introducing [rebranded Chinese model] Its more intelligent (no shit) We are doubling usage forever. Hell, next week we might even triple it.
  • @eliebakouch Elie on x
    cursor is at frontier scale, both in terms of performance and compute if composer 2.5's budget was put into a pre-train: ~6.3T total, 200B active trained on ~56T tokens if composer 3 allocates 50% of the budget to pre-training: ~500B active, 15.3T total trained on 135T tokens. [i…
  • @scaling01 @scaling01 on x
    yeah that's pretty good xAI might be able to cook with Cursor data + 10T model [image]
  • @clementdelangue Clem on x
    Very cool to see Cursor doubling down on training great models. In my opinion, ultimately all serious companies in AI will want to train models themselves, based on open-source instead of outsourcing AI to others via APIs!