Immerse yourself in the QwenWork station at the Apsara Conference!
See firsthand how this all-in-one workplace AI agent platform tackles complex tasks for individuals and enterprises across industries with ease and efficiency.
See embodied intelligence come to life at the Apsara Conference 2026!
From robots streamlining industrial automation to smart helpers for households, discover how AI bridges the gap between innovation and real-world impact.
🚀At Apsara2026, Alibaba Cloud introduced a suite of upgrates for its agentic cloud strategy, built around three pillars: Model, Harness, and Context.
These upgrades drive AI innovation at scale, spanning AI Native Cloud, Agent Native Cloud, and Context Engine to empower enterprises in creating real-world value with cutting-edge AI technologies.
At Apsara2026, Alibaba’s chip design unit, T-Head, introduced the Zhenwu V900—its latest AI training and inference processor capable of handling both high-precision model training and ultra-low-precision inference, delivering three times the performance of its predecessor.
Zhenwu AI chips are already powering 650+ customers across industries, from automobiles to LLMs and embodied intelligence.
The roadmap for next-generation Yitian CPUs was also introduced, showcasing processors tailored for agentic AI tasks and scheduled for launch in 2027.🚀
Alibaba introduced exciting new updates to the Qwen model family today at Apsara Conference 2026, including the next-gen Qwen 4 now in training, alongside advancements in multimodal models for speech, audio, and vision.
Learn more about what’s next!
Speaking at Apsara Conference 2026, Alibaba Group CEO Eddie Wu highlighted the massive frontier ahead: with Machine Thinking currently representing under 3% of human capacity, scaling toward 1,000x opens an unprecedented runway for exponential growth.
To power this future, Alibaba is building across all three pillars:
• AI Models: Advancing Recursive Self-Improvement (RSI) with plans for 5–10T parameter Qwen models driving toward ASI.
• AI Chips: Introducing the Zhenwu V900—China's most powerful AI chip, scaling up to 500,000 cards per cluster.
• AI Cloud: Expanding global data center capacity to exceed 20GW by 2032 to meet boundless demand.
At Apsara Conference 2026 in Hangzhou, Alibaba CEO Eddie Wu shared our strategic roadmap for the Machine Intelligence era:
- Scaling to 1,000x human capacity leaves an enormous growth runway ahead.
- Machines are becoming the primary force behind Thinking.
- Machine Thinking is currently <3% of all Human Thinking.
Apsara Conference 2026 starts tomorrow!
Follow a paper airplane as it soars across the expansive exhibition, offering a glimpse of four thematic pavilions dedicated to intelligence, computing power, creation, and industrial applications.
Stay tuned!
Alibaba Cloud is named a Leader in the 2026 Gartner® Magic Quadrant™ for Generative AI Model Providers and Strategic Cloud Platform Services.
This dual recognition highlights our strength in AI and cloud platforms, from Qwen models to agentive cloud infrastructure. 🚀
Qwen3.8-Omni-Flash is here, Qwen's first omni-modal model, centred on agentic capabilities and tailored for long-form audio-video understanding and agentic workflows.
#AlibabaAI#Qwen
🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities!
Native audio-video understanding, reasoning, and tool use come together in one model: understand the content, plan the task, execute with tools, and deliver the result.
Highlights: 🥳
- Audio-video intelligence that gets things done: jointly reason over what's seen and heard, and orchestrate tools across long workflows to auto-edit vlogs, translate short videos, and turn movies into recaps.
- A major leap: approaching Gemini 3.8 Flash in audio-video capabilities; +19.5 points on average in agent performance across WildClawBench-MM & UniClawBench.
- 1M-token context with agentic perception: actively explore long videos and locate key moments with higher accuracy, using 51.8% fewer tokens than static understanding on OmniVideoBench.
Video input costs are reduced by about 89% compared with Qwen3.5-Omni-Plus, making long-form audio-video understanding and agentic workflows more affordable than ever.
To help you build apps around Omni, we're also open-sourcing Qwen-MM-Plugins and Qwen-Live Harness! 🛠️
We can't wait to see what you build with Qwen3.8-Omni-Flash! 👀
- Blog: https://t.co/oM9V1TkqYF
- Qwencloud: https://t.co/Cc2I8ELAnD
- Qwen Studio: https://t.co/V7RmqMaVNZ
- API: https://t.co/lNE7fH5YUt
- Qwen-MM-Plugins: https://t.co/SnM27dDP3d
- Qwen-Live Harness: coming soon
https://t.co/iVYlGjIbdy
🚀 Apsara Conference 2026 is just around the corner!
This year, explore how technology reshapes computing, drives intelligence with cutting-edge models and agentic AI, and creates new opportunities for industries through digital transformation.
Intelligence Goes Beyond. Stay tuned for more details!
Discover how QwenWork translates complex workplace requirements into seamless, ready-to-use deliverables like spreadsheets, HTML reports, and more.
Watch the video to see how QwenWork can elevate your productivity to the next level.
Amap unveils ABot-Earth 0.7, the world’s first native 3D urban world model, powering Flying Street View 2.0.
Explore immersive 3D environments to preview theater views, navigate malls, or assess scenic terrains before visiting. Learn more in the video to see how Amap's spatial intelligence helps us better explore the world!
Utilizing years of Amap's mapping and location data, ABot-Earth 0.7 features a native 3D architecture that integrates diverse geographic and environmental information, creating a seamless, interactive digital world for immersive exploration.
Learn more: https://t.co/esoAxLm5se
#AlibabaAI
Amap Street Stars celebrates its first anniversary with a major spatial-intelligence upgrade. Flying Street View 2.0 feature lets users preview destinations as immersive and explorable 3D environments, powered by ABot-Earth 0.7, the world’s first native 3D urban world model.
With Qwen Omni, LimX Dynamics enhances real-time interaction capabilities in humanoid robots, enabling seamless processing of visual and audio inputs to handle complex user requests, leading the way in embodied intelligence.