/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← β†’ days Β· ↑ ↓ browse Β· Enter similar Β· o open

Alibaba releases the Qwen3-VL vision models, the Qwen3Guard β€œsafety moderation” models, and three closed-weight models, including Qwen3-Max with 1T+ parameters

Qwen 50.6kΒ  β€”Β  Safetensors qwen3_vl_moe Julian Nabil / Forbes Middle East : Alibaba Introduces Qwen3-Max AI Model With Over 1T Parameters Markus Kasanmascheff / WinBuzzer : Alibaba Releases Qwen3-VL Open-Source Vision Language AI Model Series Ankush Das / Analytics India Magazine : Alibaba Launches Qwen3-VL With Open Source Flagship Model Reuters : Alibaba launches Qwen3-Max AI model with more than a trillion parameters X: @alibaba_qwen : We're excited to announce the upgrade of Qwen3-Coder, and the upgraded API β€˜qwen3-coder-plus’ is now available on Alibaba Cloud Model Studio with major improvements: πŸ’» Enhanced terminal task capabilities and better performance on Terminal Bench (w/ Qwen Code / Claude Code) πŸ† [image] Ethan Mollick / @emollick : So far, Qwen3-Max seems impressive for a non-reasoning model, doing a good job at a lot of my weird tests that even some reasoners struggle with. [image] @alibaba_qwen : πŸš€ Introducing Qwen3-LiveTranslate-Flash β€” Real‑Time Multimodal Interpretation β€” See It, Hear It, Speak It! 🌐 Wide language coverage β€” Understands 18 languages & 6 dialects, speaks 10 languages. πŸ‘οΈ Vision‑Enhanced Comprehension β€” Reads lips, gestures, on‑screen text and [image] @alibaba_qwen : πŸ›‘οΈ Meet Qwen3Guard β€” the Qwen3-based safety moderation model series built for global, real-time AI safety! 🌍 Supports 119 languages and dialects βœ… 3 sizes available: 0.6B, 4B, 8B ⚑ Low-latency, Real-time streaming detection with Qwen3Guard-Stream πŸ“ Robust Full-context safety [image] Bindu Reddy / @bindureddy : The leaders in open source are far and away Qwen and DeepSeek US still lags way behind in this category For example - we still have to all the way back to Llama, if we need a fine-tune based on a US base model Awni Hannun / @awnihannun : Just for fun, here's what 32 simultaneous long-context generations with Qwen3 Next 80B looks like on an M3 Ultra. Using the new batch generation in mlx-lm. Context size for each is about 5k tokens: [video] @teortaxestex : This is what innovation looks like in ML. Lots of small things combined. Qwen3 VL is SOTA. And yet... always something subtle missing with these guys. It's the first model that has seen *a* connection and dismissed it in favor of the cat's psyop. But its vision is clear at least. [image] Omar Khattab / @lateinteraction : Sort of like calling a Qwen3-235B-A22B β€œqwen 50B” for short. Ahmad / @theahmadosman : qwen3 omni technical paper summary > qwen3-omni is one model for everything > text, vision, audio, speech, and video > beats chatgpt 4o and gemini 2.5 in reasoning and recognition > 30B model, only 3B active parameters per token > runs on consumer hardware with ease, usable [image] Andi Marafioti / @andimarafioti : Qwen3-Omni is here, and it's a huge step for omni-modal AI. πŸ”₯ Give it anything and get back text or surprisingly natural speech. Just watch the demo: it reads data from a table and then analyzes a snowboarder's balance in a video. The future is now. 🀯 [video] Junyang Lin / @justinlin610 : This is the 1st shot! For a long time people just don't have any idea about the safety work that we have invested efforts in. This time, we show you our safety guard model,Qwen3 Guard, specifically including generative guard Qwen3Guard-Gen and streaming guard model with @_akhaliq : qwen3-coder-plus is now available on Anycoder Enhanced terminal task capabilities and better performance on Terminal Bench (w/ Qwen Code / Claude Code) SWE-Bench performance up to 69.6 Safer code generation available as Qwen3-Coder-Plus-2025-09-23 [video] Elvis / @omarsar0 : Qwen3-Omni Technical Report A unified multimodal model that matches same-size Qwen text-only and vision-only baselines while pushing audio and audio-visual SOTA. Key technical details below: [image] Tianbao Xie / @tianbaox : After another half year, we are glad to bring Qwen3-VL! It's definitely the best open model you can access to start your digital agent and physical agent journey. Thanks the whole team! @shuai_bai_ @huybery @DunjieLu1219 @xuhaiya2483846 @JustinLin610 Junyang Lin / @justinlin610 : This is the 5th shot! Super crazy! We opensourced a 235B-A22B Instruct and Thinking Qwen3-VL models under Apache 2.0! Qwen3-VL, the new generation of our vision-language model, whose previous version was released a long time ago. During these days, we have conducted a lot of @sixsigmacapital : $BABA This could be Alibaba's mini chat-GPT moment. Aran Komatsuzaki / @arankomatsuzaki : RLPT: Reinforcement Learning on Pre-Training Data β€’ RL directly on pre-train data (no human labels) β€’ Next-segment reasoning objective (ASR + MSR tasks) β†’ self-supervised rewards β€’ Gains on Qwen3-4B: +3.0 MMLU, +8.1 GPQA-Diamond, +6.6 AIME24, +5.3 AIME25 [image] AshutoshShrivastava / @ai_for_success : In the last 12 hours, Qwen has released: > Qwen3Guard > Personal AI Travel Designer > Qwen3-LiveTranslate-Flash > Upgrade Qwen3-Coder > Qwen3-VL-235B-A22B > Qwen3-Max The Qwen team is crazy πŸ”₯ [image] @tryagentsea : Alibaba has: - best open weights image model (qwen image) - best open weights image editing model (qwen image edit 2509) - best sota open weights vision model (qwen3 vl) - best open weights video inpainting model (wan 2.2 animate) - one of the best foundation models (qwen3 max) Justine Chang / @justine_chang39 : Ok just did my image cropping test on Qwen3 VL It is, ON PAR, if not BETTER than Gemini 2.5 Pro for my use case 🀯 This is the FIRST non-Gemini model to be able to do this. This is really really good!! @Alibaba_Qwen @huybery @JustinLin610 [image] @reach_vb : NEW: Qwen 235B A22B Vision Language Model is OUTT! Apache 2.0 licensed and upto 1 Million context length 🀯 https://huggingface.co/... Merve / @mervenoyann : my vibe tests with Qwen3-Omni family of models > document performance with Instruct is very good 🎯 > video understanding is nice ⏯️ > Thinking performs better in English > I suggest to use Captioner if you really want audio output, other two hallucinates a bit [video] @alibaba_qwen : πŸš€ Qwen3-Max is hereβ€”no preview, just power! Qwen Chat: https://chat.qwen.ai/ Blog: https://qwen.ai/... API: https://www.alibabacloud.com/ ... We've supercharged coding & agentic skillsβ€”now Qwen3-Max-Instruct without thinking rivaling top models on SWE-Bench, Tau2-Bench, [image] Chujie Zheng / @chujiezheng : Qwen3-VL, this is what you many guys are always wanting. Enjoy 🍻 @alibaba_qwen : πŸš€ We're thrilled to unveil Qwen3-VL β€” the most powerful vision-language model in the Qwen series yet! πŸ”₯ The flagship model Qwen3-VL-235B-A22B is now open-sourced and available in both Instruct and Thinking versions: βœ… Instruct outperforms Gemini 2.5 Pro on key vision [image] Binyuan Hui / @huybery : We have released Qwen3-Max, the most powerful Qwen model to date! By continuously scaling up model size, data, and RL tasks, great things have happened. This time, coding and agent capabilities have also been significantly enhancedβ€”enjoy! Jianwei Yang / @jw2yang4ai : πŸš€Excited to see Qwen3-VL released as the new SOTA open-source vision-language model! What makes it extra special is that it's powered by DeepStack, a technique I co-developed with Lingchen, who is now a core contributor of Qwen3-VL. When Lingchen and I developed this technique Bluesky: Sung Kim / @sungkim : Chinese AI has caught up with leading U.S.-based labs.Β  Alibaba's release of Qwen3-Max places it alongside frontier AI players like Anthropic, Google, OpenAI, and xAI.Β  β€”Β  qwen.ai/blog?id=2413...Β  [images] Tim Kellogg / @timkellogg.me : Qwen3-Max: Just Scale ItΒ  β€”Β  it's now safe to say Qwen is a frontier labΒ  β€”Β  qwen.ai/blog?id=2413...Β  [image] Threads: NaveedUllah / @naveed_ullah600 : Qwen has unveiled π—€π˜„π—²π—»πŸ― 𝗧𝗧𝗦, a next gen text to speech model built for natural, expressive audio.Β  It supports English + multiple Chinese dialects (Mandarin, Cantonese, Beijing, Shanghai, Sichuan) and delivers voices that sound truly human-like. … Mastodon: Simon Willison / @simon@fedi.simonwillison.net : Plus notes on the 5 (!) new things Qwen released today, the most exciting of which is the first in their Qwen3-VL vision-LLM series, a 235B 471 GB Apache 2 licensed monster! https://simonwillison.net/... Forums: Hacker News : Qwen3-VL

Simon Willison's Weblog Simon Willison

Context & Ripple Effects

Alibaba has been extending Qwen across visual reasoning, agentic coding and multimodal inputs: Qwen2.5-VL added device-control and vision tasks, while Qwen3-Coder targeted agentic software work. This release turns those adjacent capabilities into a broader product lineup.

The mix of an Apache-licensed vision flagship, safety models and closed-weight offerings clarifies a two-track approach: wide model access alongside proprietary cloud-oriented products.

First-order effects

  • Developers can deploy or adapt the open-weight Qwen3-VL flagship, while teams needing moderation gain Qwen3Guard options for batch and streaming use across supported languages.
  • Alibaba expands its commercial model catalog with closed-weight Qwen3-Max and an upgraded Qwen3-Coder API on Alibaba Cloud Model Studio, concentrating its most proprietary coding and agent features in its service layer.

Second-order effects

  • Model buyers gain more leverage to compare open vision-language systems against proprietary alternatives, while vendors competing for enterprise workloads must match not only model capability but safety tooling and deployment choice.
  • Bundling coding, vision, translation and moderation around Alibaba Cloud Model Studio can make the cloud platform a more natural destination for customers that start with Qwen models, reinforcing distribution beyond the base model itself.

Third-order effects

  • If this split between open foundation models and closed premium services persists, competition will increasingly center on the surrounding stackβ€”safety, APIs, tools and hostingβ€”rather than on weight availability alone.
  • The addition of real-time moderation and interpretation points toward multimodal AI being sold as operational infrastructure, where reliability and governance become purchase criteria alongside benchmark performance.

The trend: AI labs are industrializing model portfolios by pairing open releases that broaden adoption with proprietary services that capture higher-value production workloads.

Discussion

  • @alibaba_qwen @alibaba_qwen on x
    We're excited to announce the upgrade of Qwen3-Coder, and the upgraded API β€˜qwen3-coder-plus’ is now available on Alibaba Cloud Model Studio with major improvements: πŸ’» Enhanced terminal task capabilities and better performance on Terminal Bench (w/ Qwen Code / Claude Code) πŸ† [ima…
  • @emollick Ethan Mollick on x
    So far, Qwen3-Max seems impressive for a non-reasoning model, doing a good job at a lot of my weird tests that even some reasoners struggle with. [image]
  • @alibaba_qwen @alibaba_qwen on x
    πŸš€ Introducing Qwen3-LiveTranslate-Flash β€” Real‑Time Multimodal Interpretation β€” See It, Hear It, Speak It! 🌐 Wide language coverage β€” Understands 18 languages & 6 dialects, speaks 10 languages. πŸ‘οΈ Vision‑Enhanced Comprehension β€” Reads lips, gestures, on‑screen text and [image]
  • @alibaba_qwen @alibaba_qwen on x
    πŸ›‘οΈ Meet Qwen3Guard β€” the Qwen3-based safety moderation model series built for global, real-time AI safety! 🌍 Supports 119 languages and dialects βœ… 3 sizes available: 0.6B, 4B, 8B ⚑ Low-latency, Real-time streaming detection with Qwen3Guard-Stream πŸ“ Robust Full-context safety [ima…
  • @bindureddy Bindu Reddy on x
    The leaders in open source are far and away Qwen and DeepSeek US still lags way behind in this category For example - we still have to all the way back to Llama, if we need a fine-tune based on a US base model
  • @awnihannun Awni Hannun on x
    Just for fun, here's what 32 simultaneous long-context generations with Qwen3 Next 80B looks like on an M3 Ultra. Using the new batch generation in mlx-lm. Context size for each is about 5k tokens: [video]
  • @teortaxestex @teortaxestex on x
    This is what innovation looks like in ML. Lots of small things combined. Qwen3 VL is SOTA. And yet... always something subtle missing with these guys. It's the first model that has seen *a* connection and dismissed it in favor of the cat's psyop. But its vision is clear at least.…
  • @lateinteraction Omar Khattab on x
    Sort of like calling a Qwen3-235B-A22B β€œqwen 50B” for short.
  • @theahmadosman Ahmad on x
    qwen3 omni technical paper summary > qwen3-omni is one model for everything > text, vision, audio, speech, and video > beats chatgpt 4o and gemini 2.5 in reasoning and recognition > 30B model, only 3B active parameters per token > runs on consumer hardware with ease, usable [imag…
  • @andimarafioti Andi Marafioti on x
    Qwen3-Omni is here, and it's a huge step for omni-modal AI. πŸ”₯ Give it anything and get back text or surprisingly natural speech. Just watch the demo: it reads data from a table and then analyzes a snowboarder's balance in a video. The future is now. 🀯 [video]
  • @justinlin610 Junyang Lin on x
    This is the 1st shot! For a long time people just don't have any idea about the safety work that we have invested efforts in. This time, we show you our safety guard model,Qwen3 Guard, specifically including generative guard Qwen3Guard-Gen and streaming guard model with
  • @_akhaliq @_akhaliq on x
    qwen3-coder-plus is now available on Anycoder Enhanced terminal task capabilities and better performance on Terminal Bench (w/ Qwen Code / Claude Code) SWE-Bench performance up to 69.6 Safer code generation available as Qwen3-Coder-Plus-2025-09-23 [video]
  • @omarsar0 Elvis on x
    Qwen3-Omni Technical Report A unified multimodal model that matches same-size Qwen text-only and vision-only baselines while pushing audio and audio-visual SOTA. Key technical details below: [image]
  • @tianbaox Tianbao Xie on x
    After another half year, we are glad to bring Qwen3-VL! It's definitely the best open model you can access to start your digital agent and physical agent journey. Thanks the whole team! @shuai_bai_ @huybery @DunjieLu1219 @xuhaiya2483846 @JustinLin610
  • @justinlin610 Junyang Lin on x
    This is the 5th shot! Super crazy! We opensourced a 235B-A22B Instruct and Thinking Qwen3-VL models under Apache 2.0! Qwen3-VL, the new generation of our vision-language model, whose previous version was released a long time ago. During these days, we have conducted a lot of
  • @sixsigmacapital @sixsigmacapital on x
    $BABA This could be Alibaba's mini chat-GPT moment.
  • @arankomatsuzaki Aran Komatsuzaki on x
    RLPT: Reinforcement Learning on Pre-Training Data β€’ RL directly on pre-train data (no human labels) β€’ Next-segment reasoning objective (ASR + MSR tasks) β†’ self-supervised rewards β€’ Gains on Qwen3-4B: +3.0 MMLU, +8.1 GPQA-Diamond, +6.6 AIME24, +5.3 AIME25 [image]
  • @ai_for_success AshutoshShrivastava on x
    In the last 12 hours, Qwen has released: > Qwen3Guard > Personal AI Travel Designer > Qwen3-LiveTranslate-Flash > Upgrade Qwen3-Coder > Qwen3-VL-235B-A22B > Qwen3-Max The Qwen team is crazy πŸ”₯ [image]
  • @tryagentsea @tryagentsea on x
    Alibaba has: - best open weights image model (qwen image) - best open weights image editing model (qwen image edit 2509) - best sota open weights vision model (qwen3 vl) - best open weights video inpainting model (wan 2.2 animate) - one of the best foundation models (qwen3 max)
  • @justine_chang39 Justine Chang on x
    Ok just did my image cropping test on Qwen3 VL It is, ON PAR, if not BETTER than Gemini 2.5 Pro for my use case 🀯 This is the FIRST non-Gemini model to be able to do this. This is really really good!! @Alibaba_Qwen @huybery @JustinLin610 [image]
  • @reach_vb @reach_vb on x
    NEW: Qwen 235B A22B Vision Language Model is OUTT! Apache 2.0 licensed and upto 1 Million context length 🀯 https://huggingface.co/...
  • @mervenoyann Merve on x
    my vibe tests with Qwen3-Omni family of models > document performance with Instruct is very good 🎯 > video understanding is nice ⏯️ > Thinking performs better in English > I suggest to use Captioner if you really want audio output, other two hallucinates a bit [video]
  • @alibaba_qwen @alibaba_qwen on x
    πŸš€ Qwen3-Max is hereβ€”no preview, just power! Qwen Chat: https://chat.qwen.ai/ Blog: https://qwen.ai/... API: https://www.alibabacloud.com/ ... We've supercharged coding & agentic skillsβ€”now Qwen3-Max-Instruct without thinking rivaling top models on SWE-Bench, Tau2-Bench, [image]
  • @chujiezheng Chujie Zheng on x
    Qwen3-VL, this is what you many guys are always wanting. Enjoy 🍻
  • @alibaba_qwen @alibaba_qwen on x
    πŸš€ We're thrilled to unveil Qwen3-VL β€” the most powerful vision-language model in the Qwen series yet! πŸ”₯ The flagship model Qwen3-VL-235B-A22B is now open-sourced and available in both Instruct and Thinking versions: βœ… Instruct outperforms Gemini 2.5 Pro on key vision [image]
  • @huybery Binyuan Hui on x
    We have released Qwen3-Max, the most powerful Qwen model to date! By continuously scaling up model size, data, and RL tasks, great things have happened. This time, coding and agent capabilities have also been significantly enhancedβ€”enjoy!
  • @jw2yang4ai Jianwei Yang on x
    πŸš€Excited to see Qwen3-VL released as the new SOTA open-source vision-language model! What makes it extra special is that it's powered by DeepStack, a technique I co-developed with Lingchen, who is now a core contributor of Qwen3-VL. When Lingchen and I developed this technique
  • @sungkim Sung Kim on bluesky
    Chinese AI has caught up with leading U.S.-based labs.Β  Alibaba's release of Qwen3-Max places it alongside frontier AI players like Anthropic, Google, OpenAI, and xAI.Β  β€”Β  qwen.ai/blog?id=2413...Β  [images]
  • @timkellogg.me Tim Kellogg on bluesky
    Qwen3-Max: Just Scale ItΒ  β€”Β  it's now safe to say Qwen is a frontier labΒ  β€”Β  qwen.ai/blog?id=2413...Β  [image]