/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Alibaba's Qwen team releases Qwen2.5-VL, a new series of AI models that can control PCs and phones, as well as perform a number of text and image analysis tasks

QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Anusuya Lahiri / Benzinga : Not Just DeepSeek - Alibaba Unveils AI Model To Rival OpenAI's Operator Markus Kasanmascheff / WinBuzzer : Alibaba Qwen Challenges OpenAI and DeepSeek with Multimodal AI Automation and 1M-Token Context Models Simon Willison / Simon Willison's Weblog : Qwen2.5 VL!  Qwen2.5 VL!  Qwen2.5 VL!  Hot on the heels of yesterday's Qwen2.5-1M … Mastodon: Alexandre Dulaunoy / @a@paperbay.org : I ran some local tests with deepseek-r1 (a distilled version using Llama and Qwen).  The reasoning output is impressive and can even be used to enhance smaller LLMs.  —  Now there is a new release of qwen which includes an improved HTML document “parsing” part and many other features. … X: @alibaba_qwen : 🎉 恭喜发财 🧧🐍 As we welcome the Chinese New Year, we're thrilled to announce the launch of Qwen2.5-VL , our latest flagship vision-language model! 🚀 💗 Qwen Chat: https://chat.qwenlm.ai/ 📖 Blog: https://qwenlm.github.io/... 🤗 Hugging Face: https://huggingface.co/... 🤖 ModelScope: [video] Binyuan Hui / @huybery : We've released the SOTA open multimodal model, Qwen2.5-VL! It shows significant improvements across various aspects compared to the previous version. I'm thrilled to have made contributions to the Agent section of Qwen2.5-VL. The journey continues—see you tomorrow! [image] @gm8xx8 : Qwen2.5-VL is a highly capable vision-language model designed for a wide range of tasks. It excels in visual understanding, device interaction, and long-form video analysis. Key Features: - Visual Understanding: Handles everything from simple objects to complex data [image] Elvis / @omarsar0 : This Qwen2.5-VL looks exciting! - strong and general vision capabilities - agentic features to support computer/phone use - long video understanding & capturing events - visualize localization - generated structured outputs Opens both base and instruct models in 3 sizes: 3B, [image] @kimmonismus : China strikes again: Qwen2.5-VL , their latest flagship vision-language model! * Visual Understanding : From flowers to complex charts, Qwen2.5-VL sees it all! * Agentic Capabilities : It's a visual agent that can reason and interact with tools like computers & phones. * Long [image] Prince Canuma / @prince_canuma : Qwen2.5-VL port to MLX update # 01 Model is loading fine just need to implement the new vision logic. 3B runs are +30 tok/s in bf16, image the quants 🔥 [image] Simon Willison / @simonw : Updated my post to highlight Qwen's fascinating new set of cookbooks https://simonwillison.net/... [image]

TechCrunch Kyle Wiggers

Discussion

  • @alibaba_qwen @alibaba_qwen on x
    🎉 恭喜发财 🧧🐍 As we welcome the Chinese New Year, we're thrilled to announce the launch of Qwen2.5-VL , our latest flagship vision-language model! 🚀 💗 Qwen Chat: https://chat.qwenlm.ai/ 📖 Blog: https://qwenlm.github.io/... 🤗 Hugging Face: https://huggingface.co/... 🤖 ModelScope: [vid…
  • @huybery Binyuan Hui on x
    We've released the SOTA open multimodal model, Qwen2.5-VL! It shows significant improvements across various aspects compared to the previous version. I'm thrilled to have made contributions to the Agent section of Qwen2.5-VL. The journey continues—see you tomorrow! [image]
  • @gm8xx8 @gm8xx8 on x
    Qwen2.5-VL is a highly capable vision-language model designed for a wide range of tasks. It excels in visual understanding, device interaction, and long-form video analysis. Key Features: - Visual Understanding: Handles everything from simple objects to complex data [image]
  • @omarsar0 Elvis on x
    This Qwen2.5-VL looks exciting! - strong and general vision capabilities - agentic features to support computer/phone use - long video understanding & capturing events - visualize localization - generated structured outputs Opens both base and instruct models in 3 sizes: 3B, [ima…
  • @kimmonismus @kimmonismus on x
    China strikes again: Qwen2.5-VL , their latest flagship vision-language model! * Visual Understanding : From flowers to complex charts, Qwen2.5-VL sees it all! * Agentic Capabilities : It's a visual agent that can reason and interact with tools like computers & phones. * Long [im…
  • @prince_canuma Prince Canuma on x
    Qwen2.5-VL port to MLX update # 01 Model is loading fine just need to implement the new vision logic. 3B runs are +30 tok/s in bf16, image the quants 🔥 [image]
  • @simonw Simon Willison on x
    Updated my post to highlight Qwen's fascinating new set of cookbooks https://simonwillison.net/... [image]