Z.ai prices its most advanced model, GLM-5.1, 8% to 17% higher than GLM-5 Turbo, joining Alibaba and Tencent in raising prices as agentic AI demand surges
Z.ai had already lifted pricing for new coding-plan subscribers as demand for its coding tools rose; this latest move extends that willingness to charge more from a product tier into the model API itself. It follows the release of GLM-5.1 as a large Mixture-of-Experts model under an MIT license, separating access to model weights from the cost of operating the service.
The subsequent product arc also points toward monetization through developer workflow products: Z.ai later introduced ZCode, an agentic development environment built around its GLM models. Higher frontier-model pricing therefore matters not only to API buyers but to the economics of tools built on top of those models.
First-order effects
Customers using GLM-5.1 face an 8%–17% higher service cost than GLM-5 Turbo, increasing the operating expense of workloads that require Z.ai's most advanced model.
Z.ai gains room to align premium-model revenue with demand for agentic workloads, after its earlier increase for new coding-plan subscribers signaled capacity and usage pressure.
Second-order effects
Developers and AI-tool vendors may shift routine tasks to cheaper models while reserving GLM-5.1 for higher-value or more difficult agentic steps, making model routing more economically consequential.
Alibaba, Tencent, and other competing providers face a clearer test of whether agentic demand supports premium pricing without pushing customers toward lower-cost alternatives or self-hosted open-weight deployments.
Third-order effects
If premium pricing persists alongside openly licensed frontier weights, competition will increasingly turn on reliable hosted inference, developer products, and the cost per completed task rather than access to weights alone.
The pattern suggests agentic AI is moving from promotional model pricing toward capacity-aware commercialization, though buyer switching and open-weight use could constrain how durable those price increases are.
The trend: Agentic AI demand is turning advanced-model access into a differentiated, capacity-constrained service whose value is measured by the economics of completed work rather than token price alone.
Zhipu raised the cost of access to its most advanced AI model by at least 8%, joining other leading Chinese AI players in trying to profit off years of research and computing investments https://www.bloomberg.com/...
we open-sourced glm-5.1 agents could do about 20 steps by the end of last year. glm-5.1 can do 1,700 rn. autonomous work time may be the most important curve after scaling laws. glm-5.1 will be the first point on that curve that the open-source community can verify with their own
Most models still break mid-task not because they're not smart enough but because they can't stay in the loop 8-hour runs start to change that this is how agents stop breaking.
Wow, GLM-5.1 beat Opus 4.6, GPT-5.4, and Gemini 3.1 Pro on SWE-Bench Pro (58.4 vs 57.3 / 57.7 / 54.2) as an open-weight MIT-licensed model! The “open-source AI vs closed-source AI” gap is still ~6 months. [image]
GLM-5.1 by @Zai_org just launched in the Text Arena, and is now the #1 open model. It outperforms the next best open model, its predecessor, GLM-5, by +11 points and +15 over Kimi K2.5 Thinking. It shows strength in: - #1 open model in Longer Query (#4 overall) - #1 open model [i…
🎉 Day-0 support for GLM-5.1 in vLLM! Congrats to @Zai_org on this next-gen flagship model built for agentic engineering, with stronger coding and sustained long-horizon task performance. Get started 👇 📖 Recipe: https://docs.vllm.ai/... [image]
Another big release: GLM-5.1! China is on fire! significant increase in evals compared to GLM-5.0 tl;dr GLM-5.1 is the new open-source agentic coding model that significantly outperforms its predecessor by sustaining long-horizon problem-solving over hundreds of iterations, [imag…
GLM-5.1 is here! Try it on OpenClaw🦞🦞🦞 ollama launch openclaw —model glm-5.1:cloud Claude Code ollama launch claude —model glm-5.1:cloud Chat with the model ollama run glm-5.1:cloud
INCREDIBLE GLM-5.1 weights are now opensource > i've had early access to the weights for the past few days > and yeah... this one matters a lot benchmarks? > SWE-Bench Pro: 58.4 > beats Opus 4.6 (57.3) > beats GPT-5.4 (57.7) > beats Gemini 3.1 Pro (54.2) let that sink in [image]
Holy moly, we thought Tuesday would be dull, but GLM-5.1 is out and it's freaking open source too 😲 What will it take to run something like this locally? A fortune? [image]
Venice delivers in <5m Available for Free to Pro users and DIEM holders in API (this model is killer for agents... first one that feels anecdotally comparable to opus imho)
damn new GLM-5.1 model crushes anthropic and openai at agentic coding and 100% open source, how does china keep getting away with this shit? - beat opus 4.6 by 6X on an open-ended coding problem - long-task horizon: 600 turns at a time + 1000s of tool calls. usually ai's just [im…
A new open model has entered the Arena! GLM-5.1 by @Zai_org is now ready for your prompts in the Text and Code Arena. Come vote and let's see how it stacks up! [image]
This is crazy! https://z.ai/ has caught up with the SOTA models in coding with its GLM-5.1 open-weight model. This is a very big deal given that billions of $ are now spent on coding tokens! Congratulations to @Zai_org team, major achievement! [image]
Vector-DB-Bench: 6x Performance Boost In high-performance database optimization, GLM-5.1 reached 21.5k QPS over 600+ iterations and 6,000+ tool calls. This is 6x the performance of a standard 50-turn session. [image]
Building a Linux Desktop from Scratch Using a self-review loop, GLM-5.1 spent 8 hours autonomously refining features, styling, and interactions to build a functional desktop environment. [video]
Introducing GLM-5.1: The Next Level of Open Source - Top-Tier Performance: #1 in open source and #3 globally across SWE-Bench Pro, Terminal-Bench, and NL2Repo. - Built for Long-Horizon Tasks: Runs autonomously for 8 hours, refining strategies through thousands of iterations. [ima…