/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

Muse Spark 1.3 with max reasoning, in limited preview for partners, scores 62 on the Artificial Analysis Intelligence Index, behind only Fable 5.1 and Opus 5

@artificialanlys

Context & Ripple Effects

Muse Spark has moved quickly through Meta’s product stack: a July public API preview for Muse Spark 1.1’s coding capabilities followed an April plan to use the model in Meta AI and shopping features. Its 1.2 benchmark result of 54 had already placed Meta alongside SpaceXAI in third among US labs.

The 1.3 result lifts that benchmark score to 62, while the highest-reasoning configuration is reserved for partners in limited preview. The release turns benchmark progress into a controlled test of where Meta can deploy more intensive reasoning.

First-order effects

  • Meta’s preview partners gain access to Muse Spark 1.3 with max reasoning, while Fable 5.1 and Opus 5 retain the two positions ahead of it on the Artificial Analysis index.
  • Meta improves its benchmark standing from the 54-point Muse Spark 1.2 result to 62, strengthening the model’s case for coding and agentic workloads in its own ecosystem.

Second-order effects

  • Fable and Opus face a closer Meta alternative in partner evaluations, making reasoning quality and the terms of access central comparison points rather than benchmark rank alone.
  • Meta can use partner feedback from the limited preview to determine whether its higher-reasoning mode is suitable for wider use across the API and Meta AI-linked products.

Third-order effects

  • The progression points to reasoning models being introduced in tiers: restricted high-capability modes first, then broader product integration once deployment economics and reliability are established.
  • If benchmark gains continue to arrive alongside gated access, AI buyers will increasingly evaluate model vendors by usable reasoning performance within their workflows, not leaderboard scores alone.

The trend: Frontier AI competition is shifting toward tiered reasoning models whose value depends on both benchmark performance and controlled access to deployment channels.

Discussion

  • @alexandr_wang Alexandr Wang on x
    1/ today we're releasing muse spark 1.3—available in muse code & the meta model api. this is our most capable model yet—frontier performance almost too cheap to meter. much stronger at agentic and coding with better usability. we think users will really notice the jump.
  • @ashvinair Ashvin Nair on x
    I recently joined the science of reasoning team at Meta! AI progress is moving faster than anyone can adjust to. I'm finding the goal of bringing personal superintelligence to everyone — AI that's positive and useful in people's lives worldwide — to be very motivating. Go try Mus…
  • @dryangsong Yang Song on x
    Muse Spark 1.3 is here! The leap from 1.0 to 1.3 is nothing short of a miracle, and it took a lot of magic to make it happen. So proud of our team of magicians who sparked this incredible leap. And we are just getting started!
  • @alexandr_wang Alexandr Wang on x
    for a single dime look at what muse spark can give you
  • @quinnypig Corey Quinn on x
    Meta casting shade at Google in AI is a playground slapfight outside a MMA championship.
  • @altryne Alex Volkov on x
    Damn, gloves are off! Unreleased Meta Muse Spark with max reasoning is beating even Fable 5, while the xhigh that is available, matches Grok 4.6 and GPT 5.6 sol on @ArtificialAnlys This is quite the statement from @AIatMeta 🔥 Busy weeks ahead of us!
  • @alexandr_wang Alexandr Wang on x
    muse spark 1.3 + muse code evals competitively with Claude code + opus 5 and Claude code + fable 5
  • @bnjmn_marie Benjamin Marie on x
    Google was ahead only a few hours. And Meta will release the weights.
  • @xeophon Florian Brand on x
    Gemini 3.8 held a spot at the pareto frontier for *checks notes* 3.5 hours
  • @deryatr_ Derya Unutmaz on x
    I wanted to try something quick with Muse Spark 1.3, so I had it build this pirate platformer. It was incredibly fast and worked flawlessly at one-shot! Now I'm enhancing the artwork, sprites and adding new levels 😊This is such a fun coding model, it just works, fast and cheap!
  • @alexandr_wang Alexandr Wang on x
    i really hate to say it, but... gemini who? 🏎️💨
  • @hesamation @hesamation on x
    Muse Spark 1.3 is now the top non-Anthropic model on Artificial Analysis Intelligence Index: > same score as Fable 5 > with 12x cheaper output tokens ($4.25/M vs $50/M) > just 1 point behind Opus 5, 4 points behind Fable 5.1 from YESTERDAY.  I'm especially curious about cost per …
  • @metafordevs @metafordevs on x
    Muse Spark 1.3 is now available in Muse Code and Meta Model API. It's tuned for the agentic builds developers actually ship, including long-running, multi-agent workflows. 🧵👇(1/4) [video]
  • @alexandr_wang Alexandr Wang on x
    underrated result of muse spark 1.3 try it out with muse code: curl -fsSL https://dev.meta.ai/... | bash
  • @aiatmeta @aiatmeta on x
    We're excited to release Muse Spark 1.3 with improved performance on agentic and coding tasks, and a focus on real-world usability. Key capabilities: → Sustains longer-horizon work across multiple workflows in a single thread → More actively collaborates with users: it asks clari…
  • @edwardsun0909 Zhiqing Sun on x
    We posttrained avocado again and enabled better test-time scaling. It's significantly better than its predecessor in all capabilities!
  • @dandr1s Dan on x
    Meta's Muse Spark 1.3 just scored 62 on Artificial Analysis' Intelligence Index. That ties Claude Fable 5, but Muse costs 8x less for input and nearly 12x less for output. It also scores above GPT-5.6 Sol, Grok 4.6, Kimi K3, and Gemini 3.8 Flash. Meta is suddenly in the top tier …
  • @openrouter @openrouter on x
    Muse Spark 1.3 is now available on OpenRouter! Built for long-running agentic, multi-agent, and coding workflows. It tracks what it learns, handles conflicting inputs, and asks for clarification or confirmation when needed. Use it now: https://openrouter.ai/...
  • @mihawkxxxxx @mihawkxxxxx on x
    @DanDr1s Incredible Minecraft from muse spark 1.3 cost 10 cents
  • @nateberkopec Nate Berkopec on x
    Muse Spark 1.3 xhigh is frontier for time/task AND cost/task. We have a new frontier human-in-the-loop coding model. Gemini 3.8-flash is slightly cheaper but also slower than 3.7, so no changes there.
  • @kimmonismus @kimmonismus on x
    It's worth really grasping this: in a very short time, the race between OpenAI and Anthropic has turned into a contest involving OpenAI, Anthropic, xAI, and Meta, and Google is back in the mix, too.  They are all on (mostly) equal footing, with little difference between them even…
  • @kimmonismus @kimmonismus on x
    Today meta chose war with google. But hey, let them fight it out. That just means development will accelerate even further. By the way: I can't imagine OpenAI waiting long to reclaim the top spot in DeepSWE.
  • @mattdeitke Matt Deitke on x
    The jump for Muse Spark 1.3 on AA Index is quite strong! Only behind recent Anthropic models at this point. Much stronger models coming soon! 🍉🥳
  • @zephyr_z9 @zephyr_z9 on x
    Very impressive
  • @angaisb_ Angel on x
    Not Meta mogging Gemini 3.8 Flash the same day lmao [image]
  • @haider1 Haider on x
    how tf Meta is moving this insanely fast needs to be studied they went from looking weirdly behind in the AI race to suddenly shipping at a pace that feels completely different
  • @chrisgpt Chris on x
    We're starting to see the fruits of Metas massive spend. So happy to see FAIR get its footing.
  • @spac89 @spac89 on x
    How the hell is Muse Spark 1.3 Max better than Fable 5? I genuinely don't think anyone saw this coming..
  • @udiwertheimer Udi Wertheimer on x
    what is even the point of anthropic anymore
  • @artificialanlys @artificialanlys on x
    Muse Spark 1.3 (max) scores 52% on Tau3-Bench Banking, the new #1 performer on this evaluation. Muse Spark 1.3 (xhigh) scores 47%, level with Claude Fable 5.1 (max, 47%) and GLM-5.3-Flash (47%) and behind its sibling in addition to e.g. Qwen3.8 Max (51%), Grok 4.6 (high, 51%), an…
  • @artificialanlys @artificialanlys on x
    The gains for Muse Spark 1.3 (xhigh) and Muse Spark 1.3 (max) over Muse Spark 1.2 on the Artificial Analysis Intelligence Index are concentrated in agentic evaluations: GDPval-AA v2 +94 and +139 Elo (1615 to 1709 and 1754), Terminal-Bench 2.1 +5 and +6 points (80% to 85% and 86%)…
  • @artificialanlys @artificialanlys on x
    Muse Spark 1.3 (xhigh) is the most cost-efficient model at its intelligence level: $0.55 per Intelligence Index task at Meta's unchanged $1.25/$4.25 per 1M token pricing.  No model scoring 59 or above costs less per task.  The nearest are Gemini 3.8 Flash (high, 59, $0.58), GPT-5…
  • @rihardjarc Rihard Jarc on x
    Wow, $META full Scorched-Earth Strategy with this AI model pricing. Nobody wants to have Zuck as a competitor.
  • @finkd Mark Zuckerberg on x
    Mark Zuckerberg says Meta's Watermelon model and Muse Spark open weights are “coming soon”
  • Brian Gamido Brian Gamido on linkedin
    We just released Muse Spark 1.3.  —  It's Meta's biggest jump so far on AI performance, especially on agentic and coding tasks …