/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Z.ai prices its most advanced model, GLM-5.1, 8% to 17% higher than GLM-5 Turbo, joining Alibaba and Tencent in raising prices as agentic AI demand surges

Luz Ding /Bloomberg:

Bloomberg Luz Ding

Context & Ripple Effects

Z.ai had already lifted pricing for new coding-plan subscribers as demand for its coding tools rose; this latest move extends that willingness to charge more from a product tier into the model API itself. It follows the release of GLM-5.1 as a large Mixture-of-Experts model under an MIT license, separating access to model weights from the cost of operating the service.

The subsequent product arc also points toward monetization through developer workflow products: Z.ai later introduced ZCode, an agentic development environment built around its GLM models. Higher frontier-model pricing therefore matters not only to API buyers but to the economics of tools built on top of those models.

First-order effects

  • Customers using GLM-5.1 face an 8%–17% higher service cost than GLM-5 Turbo, increasing the operating expense of workloads that require Z.ai's most advanced model.
  • Z.ai gains room to align premium-model revenue with demand for agentic workloads, after its earlier increase for new coding-plan subscribers signaled capacity and usage pressure.

Second-order effects

  • Developers and AI-tool vendors may shift routine tasks to cheaper models while reserving GLM-5.1 for higher-value or more difficult agentic steps, making model routing more economically consequential.
  • Alibaba, Tencent, and other competing providers face a clearer test of whether agentic demand supports premium pricing without pushing customers toward lower-cost alternatives or self-hosted open-weight deployments.

Third-order effects

  • If premium pricing persists alongside openly licensed frontier weights, competition will increasingly turn on reliable hosted inference, developer products, and the cost per completed task rather than access to weights alone.
  • The pattern suggests agentic AI is moving from promotional model pricing toward capacity-aware commercialization, though buyer switching and open-weight use could constrain how durable those price increases are.

The trend: Agentic AI demand is turning advanced-model access into a differentiated, capacity-constrained service whose value is measured by the economics of completed work rather than token price alone.

Discussion

  • @business @business on x
    Zhipu raised the cost of access to its most advanced AI model by at least 8%, joining other leading Chinese AI players in trying to profit off years of research and computing investments https://www.bloomberg.com/...
  • @louszbd Lou on x
    we open-sourced glm-5.1 agents could do about 20 steps by the end of last year. glm-5.1 can do 1,700 rn. autonomous work time may be the most important curve after scaling laws. glm-5.1 will be the first point on that curve that the open-source community can verify with their own
  • @simonw Simon Willison on x
    754B parameters, 1.51TB on Hugging Face
  • @iterintellectus Vittorio on x
    1) what [image]
  • @zaiforstartups @zaiforstartups on x
    Most models still break mid-task not because they're not smart enough but because they can't stay in the loop 8-hour runs start to change that this is how agents stop breaking.
  • @zixuanli_ Zixuan Li on x
    Going quiet on X for a few days usually means something big is coming
  • @yuchenj_uw Yuchen Jin on x
    Wow, GLM-5.1 beat Opus 4.6, GPT-5.4, and Gemini 3.1 Pro on SWE-Bench Pro (58.4 vs 57.3 / 57.7 / 54.2) as an open-weight MIT-licensed model! The “open-source AI vs closed-source AI” gap is still ~6 months. [image]
  • @arena @arena on x
    GLM-5.1 by @Zai_org just launched in the Text Arena, and is now the #1 open model. It outperforms the next best open model, its predecessor, GLM-5, by +11 points and +15 over Kimi K2.5 Thinking. It shows strength in: - #1 open model in Longer Query (#4 overall) - #1 open model [i…
  • @vllm_project @vllm_project on x
    🎉 Day-0 support for GLM-5.1 in vLLM! Congrats to @Zai_org on this next-gen flagship model built for agentic engineering, with stronger coding and sustained long-horizon task performance. Get started 👇 📖 Recipe: https://docs.vllm.ai/... [image]
  • @kimmonismus @kimmonismus on x
    Another big release: GLM-5.1! China is on fire! significant increase in evals compared to GLM-5.0 tl;dr GLM-5.1 is the new open-source agentic coding model that significantly outperforms its predecessor by sustaining long-horizon problem-solving over hundreds of iterations, [imag…
  • @ollama @ollama on x
    GLM-5.1 is here! Try it on OpenClaw🦞🦞🦞 ollama launch openclaw —model glm-5.1:cloud Claude Code ollama launch claude —model glm-5.1:cloud Chat with the model ollama run glm-5.1:cloud
  • @theahmadosman Ahmad on x
    INCREDIBLE GLM-5.1 weights are now opensource > i've had early access to the weights for the past few days > and yeah... this one matters a lot benchmarks? > SWE-Bench Pro: 58.4 > beats Opus 4.6 (57.3) > beats GPT-5.4 (57.7) > beats Gemini 3.1 Pro (54.2) let that sink in [image]
  • @zephyr_z9 @zephyr_z9 on x
    BIG [image]
  • @ai_for_success AshutoshShrivastava on x
    Holy moly, we thought Tuesday would be dull, but GLM-5.1 is out and it's freaking open source too 😲 What will it take to run something like this locally? A fortune? [image]
  • @erikvoorhees Erik Voorhees on x
    Venice delivers in <5m Available for Free to Pro users and DIEM holders in API (this model is killer for agents... first one that feels anecdotally comparable to opus imho)
  • @cryptopunk7213 @cryptopunk7213 on x
    damn new GLM-5.1 model crushes anthropic and openai at agentic coding and 100% open source, how does china keep getting away with this shit? - beat opus 4.6 by 6X on an open-ended coding problem - long-task horizon: 600 turns at a time + 1000s of tool calls. usually ai's just [im…
  • @arena @arena on x
    A new open model has entered the Arena! GLM-5.1 by @Zai_org is now ready for your prompts in the Text and Code Arena. Come vote and let's see how it stacks up! [image]
  • @deryatr_ Derya Unutmaz on x
    This is crazy! https://z.ai/ has caught up with the SOTA models in coding with its GLM-5.1 open-weight model. This is a very big deal given that billions of $ are now spent on coding tokens! Congratulations to @Zai_org team, major achievement! [image]
  • @clementdelangue Clem on x
    The best performing model on SWE-Bench Pro is open-source on @huggingface! Welcome GLM 5.1! https://huggingface.co/... [image]
  • @eliebakouch Elie on x
    GLM-5.1 sota on SWE Bench Pro 😮 [image]
  • @zai_org @zai_org on x
    Vector-DB-Bench: 6x Performance Boost In high-performance database optimization, GLM-5.1 reached 21.5k QPS over 600+ iterations and 6,000+ tool calls. This is 6x the performance of a standard 50-turn session. [image]
  • @zai_org @zai_org on x
    Building a Linux Desktop from Scratch Using a self-review loop, GLM-5.1 spent 8 hours autonomously refining features, styling, and interactions to build a functional desktop environment. [video]
  • @zai_org @zai_org on x
    SOTA on SWE-Bench Pro (58.4): GLM-5.1 delivers significant leaps in coding and agentic performance. [image]
  • @zai_org @zai_org on x
    Introducing GLM-5.1: The Next Level of Open Source - Top-Tier Performance: #1 in open source and #3 globally across SWE-Bench Pro, Terminal-Bench, and NL2Repo. - Built for Long-Horizon Tasks: Runs autonomously for 8 hours, refining strategies through thousands of iterations. [ima…
  • r/LocalLLaMA r on reddit
    GLM-5.1