Alibaba Cloud Community (via Future Tools)

Alibaba Unveils Qwen3.8-Max: Its Largest and Most Capable Flagship Model to Date

Brief

Qwen3.8-Max is Alibaba’s new flagship multimodal foundation model (announced August 3, 2026) that combines a Sparse Mixture-of-Experts architecture with a hybrid attention mechanism to deliver large-scale capability with inference efficiency: 2.4 trillion total parameters but only ~95 billion activated per inference, and a context window up to one million tokens. Built on Qwen 3.5, it achieves competitive leaderboard placements (5th Text Arena, 2nd Vision Arena, 4th Frontend Code Arena), outperformed humans in the WWW2025 multimodal intent task, and autonomously completed a 16-day engineering project that produced the open-source oh-my-cli agent framework. Accessible via Alibaba Cloud Model Studio and QwenWork (weights slated for release the week after the announcement), Qwen3.8-Max emphasizes long-horizon, visual, and closed-loop agentic workflows—ingesting hundred-page docs, full TV series or ~100-hour livestreams, reconstructing apps in black-box RecreationBench tests, and converting 2D plans, screenshots, or raw footage into production-ready outputs.

Why it matters

Alibaba announced Qwen3.8-Max on August 3, 2026: a multimodal foundation model with 2.4 trillion parameters and an extended context window up to 1,000,000 tokens; model weights were scheduled for public release the week after the announcement and the model is available via Alibaba Cloud Model Studio APIs and QwenWork.

Key details

  • Architecturally, Qwen3.8-Max uses a Sparse Mixture-of-Experts (MoE) design with a hybrid attention mechanism that keeps inference efficient by activating only ~95 billion parameters despite the 2.4T total size.
  • In benchmarks and tasks, Qwen3.8-Max ranked 5th in Text Arena, 2nd in Vision Arena, and 4th in Frontend Code Arena; it outperformed humans in the WWW2025 Multimodal Dialogue Intent Recognition Challenge and autonomously executed a 16-day software engineering project that produced the open-sourced 'oh-my-cli'.
  • The model supports large-scale visual and long-horizon workloads (e.g., ingesting hundred-page documents, full TV series or ~100-hour livestreams), demonstrated black-box app reconstruction on the new RecreationBench, and claims end-to-end autonomous planning and adaptive learning across RL-enabled agent frameworks.
Cleaned source text

Blog

Events

Webinars

Tutorials

Forum

Qwen3.8-Max exhibits advanced capabilities in coding, real-life work, research, and long-horizon tasks.

Hangzhou, China, August 3, 2026 - Alibaba officially announced the launch of Qwen3.8-Max, the most powerful model in its Qwen series to date. Boasting 2.4 trillion parameters and supporting a context window of up to 1 million tokens, Qwen3.8-Max ranks fifth in Text Arena and second in Vision Arena, while exhibiting advanced capabilities in coding, real-life work, research, and long-horizon tasks. Operating as a multimodal foundation, Qwen3.8-Max also supports visual intelligence.

The model is now accessible via APIs on the Alibaba Cloud Model Studio for global developers, with model weights scheduled for release next week. It can be also experienced on QwenWork, Alibaba’s latest all-in-one workplace AI agent platform.

Qwen3.8-Max ranks fifth in Text Arena and second in Vision Arena

Built upon the robust foundation of Qwen 3.5, Qwen3.8-Max features a cutting-edge Sparse Mixture-of-Experts (MoE) architecture combined with a hybrid attention mechanism. This design achieves a critical balance between massive scale and inference efficiency. Despite its total size of 2.4 trillion parameters, Qwen3.8-Max activates just 95 billion parameters. This innovative framework allows the model to deliver frontier-level intelligence for complex tasks while significantly reducing computational costs and latency compared to traditional dense models of similar scale.

Ranked fourth in Frontend Code Arena, Qwen3.8-Max demonstrates exceptional proficiency in autonomous coding and long-horizon execution, operating independently over extended periods without human intervention. In internal testing, the model autonomously executed a real-world software engineering project over a 16-day period. Tasked with creating a self-evolving agent framework from scratch, Qwen3.8-Max established an engineering loop that synthesized user feedback, community best practices, and self-test data. By continuously iterating through code generation, testing, previewing, and log analysis, it produced "oh-my-cli", a self-evolving agent framework that has been fully open-sourced on GitHub.

Beyond coding automation, Qwen3.8-Max drives genuine innovation in highly specialized fields. For example, the model not only reproduced published research experiments but also engineered novel methodologies that outperformed the original papers. Proving its human-level competence, Qwen3.8-Max outperformed human participants in the WWW2025 Multimodal Dialogue Intent Recognition Challenge, showcasing an unmatched ability to analyze customer-service transcripts and accurately interpret complex user demands.

Additionally, Qwen3.8-Max is engineered to manage intricate, real-world workloads, spanning application design, legal document review, sports analytics, financial research, culinary concept development, rehabilitation progress visualization, and architectural 3D modeling. By jointly scaling reinforcement learning environments and compute, the model significantly enhances general operational competence across mainstream agent frameworks. When addressing highly complex, multi-constraint, and long-horizon challenges, Qwen3.8-Max also exhibits elite system-level autonomous planning and end-to-end, closed-loop adaptive learning capabilities.

Operating as a multimodal foundation, Qwen3.8-Max supports visual intelligence, transforming static inputs into dynamic, interactive knowledge structures. The model can seamlessly ingest hundred-page documents, full television series, or 100-hour livestreams, converting them into searchable, interactive knowledge bases.

Qwen3.8-Max excels at continuous execution and creation based on real-time visual feedback. Its versatile capabilities include editing raw personal footage into professional vlogs, generating immersive educational animations from text prompts, reconstructing complete frontend web projects from a single user-interface screenshot, transforming 2D floor plans into detailed 3D interior visualizations, and building interactive games directly from natural language requests.

To demonstrate this, the team introduced RecreationBench, a long-horizon application-recreation benchmark. Operating entirely in a black-box environment with no internet access or source code visibility, Qwen3.8-Max autonomously reconstructed the applications from scratch by evaluating live applications purely through interaction and feedback. It shows by relying entirely on interactive evaluation and visual feedback, the model demonstrated frontier hybrid agent capabilities, advancing visual coding through iterative development.

Alibaba Launches "QwenWork," an All-in-One Workplace AI Agent Platform

Alibaba Cloud Community1,493 posts | 508 followers Follow

Qwen3.8-Max: A New Bar for Coding and CoworkAlibaba Cloud Community - August 3, 2026

Alibaba Cloud Unveils Strategic Roadmaps for the Next Generation AI InnovationsAlibaba Cloud Community - September 27, 2025

Model Studio Token Plan for Individual One Subscription for Every AI Model, Up to 3x More ValueAlibaba Cloud Community - August 4, 2026

Alibaba Unveils New AI Chip, Flagship Model, and Rebuilt Cloud Stack AI for Agentic EraAlibaba Cloud Community - May 20, 2026

Alibaba Cloud Unveils New AI Models and Revamped Infrastructure for AI ComputingAlibaba Cloud Community - September 19, 2024

Alibaba Unveils Qwen Glasses at MWC Barcelona, Accelerating AI Hardware AmbitionsAlibaba Cloud Community - March 4, 2026

Token PlanBuild more, spend less. One plan, every modality. Learn More

Alibaba Cloud Model StudioA one-stop generative AI platform to build intelligent applications that understand your business, based on Qwen model series such as Qwen-Max and other popular models Learn More

QwenFull-range, open-source, multimodal, and multi-functional Learn More

Offline Visual Intelligence Software PackagesOffline SDKs for visual production, such as image segmentation, video segmentation, and character recognition, based on deep learning technologies developed by Alibaba Cloud. Learn More

More Posts

by Alibaba Cloud Community

Alibaba Cloud Smart Studio Self-Service Edition Now Live Internationally

Alibaba Unveils Wan3.0 with Twice as Long Video Outputs from a Richer Variety of Inputs

Introducing Qoder Code Security: Security From the First Line of Code

Terminal Sandbox: Infrastructure for the Agent Harness

How We Used Qoder to Let an Agent Iterate on Itself: Computer Use as an Example

Proactive Memory Agent: The Agent's Second Brain

What It Actually Takes to Run Qwen3.8-27B Locally

Model Studio Token Plan for Individual One Subscription for Every AI Model, Up to 3x More Value

Qwen3.8-Max: A New Bar for Coding and Cowork