| Field | Value | ||
|---|---|---|---|
| Type | AI research team inside a hyperscaler | ||
| Parent | Alibaba Cloud (Alibaba Group) | ||
| Country | China | ||
| Flagship model family | [[qwen3 | Qwen]] (Qwen1 → Qwen1.5 → Qwen2 → Qwen2.5 → Qwen3 → Qwen3.7-Max → [[models/qwen38-max | Qwen3.8-Max]]) |
| Licensing posture | Apache 2.0 for Qwen3 weights; proprietary, API-only for Qwen3.7-Max; open weights promised for Qwen3.8-Max |
The Qwen team is Alibaba Cloud's foundation-model group. It has shipped a multi-generation family of open-weight large language models and is, together with DeepSeek, one of the two most prolific Chinese producers of frontier-competitive open-weight models. Its May 2025 Qwen3 release applied an Apache 2.0 license to a 235B-parameter mixture-of-experts (MoE) flagship, one of the most permissive licenses applied to a frontier-scale model to that point (Source: Qwen3 Technical Report). In May 2026 Alibaba departed from that posture for its top tier, releasing the agentic flagship Qwen3.7-Max as a proprietary, API-only model.
Snapshot
Models
| Date | Model | Type / license | Key facts | Source | |
|---|---|---|---|---|---|
| 2026-07-19 | [[models/qwen38-max | Qwen3.8-Max-Preview]] | Preview; open weights promised "soon" (no license or model card published) | 2.4T-parameter flagship; Alibaba describes it as "second only to Fable 5"; unveiled at WAIC Shanghai; preview live on Token Plan, Qoder, and QoderWork at 10% of standard price | (Source: bloomberg.com; the-decoder.com) |
| 2026-05-21 | Qwen3.7-Max | Proprietary, API-only | Long-horizon agentic flagship; supports external agent harnesses including Anthropic's [[companies/anthropic | Claude Code]] | (Source: venturebeat.com) |
| 2025-05 | [[qwen3 | Qwen3]] | Open weights, Apache 2.0 | 8 models, 0.6B–235B; 235B MoE flagship; unified thinking/non-thinking modes with adjustable thinking budget; 119 languages | (Source: Qwen3 Technical Report) |
| (prior) | Qwen2.5 | Open weights | 29-language support; specialist variants Qwen2.5-VL (vision-language), Qwen2.5-Math, Qwen2.5-Coder | (Source: Qwen3 Technical Report) | |
| (prior) | Qwen2 | Open weights | Expanded scale and multilingual support | (Source: Qwen3 Technical Report) | |
| (prior) | Qwen / Qwen1.5 | Open weights | Early open-weight Chinese LLMs; dense architectures | (Source: Qwen3 Technical Report) |
Qwen3.7-Max claimed performance (Alibaba figures)
| Metric | Value | Source |
|---|---|---|
| Continuous autonomous execution | ~35 hours | (Source: venturebeat.com) |
| Tool calls in that run | 1,158 | (Source: venturebeat.com) |
| Kernel evaluations in that run | 432 | (Source: venturebeat.com) |
| Apex Math Reasoning benchmark | 44.5 (vs. Claude Opus 4.6 Max 34.5) | (Source: venturebeat.com) |
The ~35-hour, 1,158-tool-call autonomous-run figures are Alibaba's own claims and are not yet independently verified; confidence on these figures is medium.
Overview and relationship to Alibaba Cloud
The Qwen team sits inside Alibaba Cloud and serves both as an AI capability for Alibaba's own products and as a model-as-a-service offering. Its open-weight releases complement, rather than replace, Alibaba Cloud's hosted API. This posture differs from OpenAI and Anthropic (closed weights, API-only) and from Meta (open weights, but not a public-cloud customer-serving posture).
Product family
The Qwen line runs from the early open-weight Qwen and Qwen1.5 dense models through Qwen2 (expanded scale and multilingual support) to Qwen2.5, which added 29-language support and spawned specialist variants Qwen2.5-VL (vision-language), Qwen2.5-Math, and Qwen2.5-Coder (Source: Qwen3 Technical Report). The May 2025 Qwen3 release comprised 8 models from 0.6B to 235B parameters, with unified thinking/non-thinking modes, an adjustable thinking budget, and 119-language coverage, all under Apache 2.0 (Source: Qwen3 Technical Report). In May 2026 Alibaba added Qwen3.7-Max, a long-horizon agentic flagship released proprietary and API-only.
Specialist Qwen2.5 variants were used as synthetic-data generators for Qwen3 pre-training: Qwen2.5-VL for PDF text extraction, and Qwen2.5-Math and Qwen2.5-Coder for domain-specific synthetic data (Source: Qwen3 Technical Report). This is a self-improvement loop inside the Qwen family, comparable to the R1→V3 distillation inside DeepSeek (see DeepSeek-V3 Technical Report).
Qwen3.7-Max
Alibaba released Qwen3.7-Max on May 21, 2026 as an agentic, long-horizon flagship. Alibaba says the model ran for roughly 35 hours of continuous autonomous execution, comprising 1,158 tool calls and 432 kernel evaluations, optimizing code on hardware unseen during training, and scored 44.5 on the Apex Math Reasoning benchmark against Claude Opus 4.6 Max's 34.5 (Source: venturebeat.com). The model is designed to support external agent harnesses, including Anthropic's Claude Code (Source: venturebeat.com).
Qwen3.7-Max was released proprietary and API-only, departing from Alibaba's prior open-weight releases such as the Apache 2.0 Qwen3 family. The closed release of the flagship tier indicates Alibaba is treating its most capable agentic model as a commercial API asset rather than an open-weight release, while the open Qwen3 family persists at lower tiers. This bears on the question, discussed in Open-Source AI / Open-Weight Models and China and the US Are Running Different AI Races, of whether Chinese labs are tilting toward closed releases (see Licensing and policy posture, below).
Qwen3.8-Max preview
On July 19, 2026, the Qwen team unveiled Qwen3.8-Max-Preview, a 2.4-trillion-parameter flagship it describes as "second only to Fable 5," at the World Artificial Intelligence Conference in Shanghai, two days after Moonshot's Kimi K3 release (Source: bloomberg.com; the-decoder.com). Open weights are promised "soon" — a return to open release at the flagship tier after the closed Qwen3.7-Max — and the preview is live on Alibaba's Token Plan, Qoder, and QoderWork at 10% of standard price, though no license, model card, or benchmarks had been published at the preview announcement (Source: the-decoder.com). Alibaba's shares rose as much as 5.4% on July 20 (Source: bloomberg.com; Who's Afraid of Chinese Models? (Ben Thompson, Stratechery, July 2026)). The announcement followed Xi Jinping's July 18 speech urging China to "encourage open source, openness, collaboration and sharing" in AI, which Stratechery's Ben Thompson read as preceding Alibaba's return to open-weight releases (Source: english.scio.gov.cn; Who's Afraid of Chinese Models? (Ben Thompson, Stratechery, July 2026)). See Chinese AI Policy.
Qwen 3.8 27B
The Qwen team released Qwen 3.8 27B on Friday, August 14, 2026: an Apache 2.0-licensed, 27-billion-parameter vision-capable model that ships defaulting to the xhigh reasoning-effort setting. Testing the 17GB Q4_K_M build on an M5 Max MacBook Pro and an NVIDIA DGX Spark, Simon Willison reported on August 16 that a single SVG drawing took 21 minutes and 22,276 reasoning tokens to produce 3,223 output tokens at that default, against 3,715 tokens in 137 seconds with reasoning turned off, and measured 15 to 30 tokens per second in LM Studio (Source: simonwillison.net). The release continues the open-weight line at the mid tier alongside the flagship Qwen3.8-Max.
Ecosystem scale and download figures
Alibaba said in an emailed statement reported on August 15, 2026 that Qwen has open-sourced more than 460 models, that its ecosystem has spawned more than 300,000 derivatives, and that its open-weight models drew more than 3 billion global downloads in the past six months, against 418 million for Google and 227 million for Meta in 2026 on Hugging Face's figures (Source: bloomberg.com).
Hugging Face's own State of Open Models report of August 14, 2026 counted 151,448 Qwen-based derivatives on its Hub, 2.6 times Meta's total footprint and ahead of Google at 82,506 (Source: huggingface.co). The two derivative counts differ by roughly a factor of two and neither source reconciles them: Hugging Face's is Hub-only and independently measured, while Alibaba's is ecosystem-wide and self-reported.
Chip software stack
Alibaba open-sourced its chip software stack in a move reported as targeting Nvidia's dominant CUDA ecosystem, per reporting circulating July 19, 2026 (Source: scmp.com). See Semiconductor Supply Chain.
Licensing and policy posture
Qwen3 is licensed under Apache 2.0, which permits commercial use, modification, and redistribution (Source: Qwen3 Technical Report). Its 119-language coverage and permissive license make it relevant to AI Sovereignty debates, since non-aligned states and multilingual deployments can use Qwen3 without US API dependency.
Alibaba's frontier-scale Apache 2.0 release runs against the framing in Open-Source AI / Open-Weight Models and the Epoch AI account (Source: epochai.substack.com) that "Alibaba has recently tilted back toward closed releases"; as of May 2025 Alibaba was still shipping its frontier-scale flagship open. The May 2026 proprietary, API-only release of Qwen3.7-Max moves in the opposite direction at the flagship tier, while the open Qwen3 family continues at lower tiers, leaving Alibaba's posture bifurcated between an open mid-tier and a closed flagship.
In the broader US-China competition, Alibaba and DeepSeek together anchor the pattern often summarized as "China ships open, US ships closed" at the frontier (China and the US Are Running Different AI Races), with Qwen3's Apache 2.0 release at 235B as a strong data point and the Qwen family serving as a counterexample to claims that Chinese labs are moving uniformly toward closed releases (Open-Source AI / Open-Weight Models). Strong-to-weak distillation inside the Qwen family parallels the efficiency strategies documented at DeepSeek (Distillation, Fast-Follow Problem). The combination of Apache 2.0 licensing and 119-language coverage positions Qwen3 as a substrate for national and regional AI deployments outside the US or China (AI Sovereignty).
A further shift in the commercial terms attached to the open tier became public on August 7, 2026, when two people familiar with the plans said Alibaba intends to ask major users of the next version of its Qwen open-source model for a share of the revenue they make from it. The measure was described as due to be implemented the following week and as applying to Qwen3.8-Max, whose learned settings are available for download. To that point Alibaba had charged developers for use of its models hosted on its own cloud platform while allowing most of its open-source offerings to run in customers' own data centres without payment. The rate had not been settled, with discussions ongoing (Source: reuters.com). The approach follows the licence Moonshot AI applies to Kimi K3, and would narrow the practical difference between an open-weight release and a licensed one for the largest downstream users; see Open-Weight Frontier Models.
On June 24, 2026 Anthropic publicly accused Alibaba of "brazenly" and "illicitly" attempting to extract its AI capabilities, in a June 10 letter to the U.S. Senate Committee on Banking, Housing, and Urban Affairs (Senators Tim Scott and Elizabeth Warren) first reported by Bloomberg. Anthropic called it "the largest known distillation attack on Anthropic to date," alleging operators it tied to Alibaba and its AI lab conducted 28.8 million model exchanges through roughly 25,000 fraudulent accounts between April 22 and June 5, 2026. Alibaba had not publicly responded as of the reports; the accusation extends the adversarial-distillation framing previously directed at DeepSeek, Moonshot, and MiniMax (Source: cnbc.com; bloomberg.com).
Alibaba banned employees from using Anthropic's Claude Code and ordered Claude models removed from work computers, according to reporting on July 3, 2026, citing alleged backdoor risks and directing staff to its own Qoder coding platform. The ban, which takes effect July 10, 2026 and classifies Claude Code as high-risk software, followed the June distillation accusation and developer reports that Claude Code inspected user environments — behavior an Anthropic employee described on June 30 as an anti-abuse "experiment we launched in March" (Source: reuters.com; theinformation.com). Anthropic's Thariq Shihipar acknowledged on July 4 that the March experiment could identify Chinese users, saying it targeted reseller abuse and distillation and that stronger mitigations had since been deployed (Source: techcrunch.com).
Alibaba and Tencent backed Kuaishou's Kling AI video unit in a $2.8 billion fundraise announced July 3, 2026 (Source: reuters.com).
Alibaba was among the firms — with ByteDance and Z.ai — that Chinese authorities met over the month before July 7, 2026 about restricting overseas access to China's most advanced AI models, including unreleased ones; the Ministry of Commerce-led discussions also covered criminalizing theft or leaks of proprietary AI technology and new rules on funding domestic AI startups (Source: reuters.com). See Export Controls (AI).
Enterprise adoption and the data-security debate
On July 15, 2026, Alibaba confirmed that its Qwen model will power Apple Intelligence in China after the Cyberspace Administration of China added Apple's AI services to its approved list; Alibaba's U.S.-listed shares rose 4% in premarket trading on the confirmation (Source: cnbc.com). The integration concludes the Apple-Alibaba China AI rollout first reported in 2025 and delayed amid US-China trade tensions.
In a Bloomberg TV interview on May 20, 2026, Airbnb CEO Brian Chesky defended Airbnb's use of Qwen for its customer-service chatbot and contested the framing that Chinese-developed open-weight models pose data-security risks, stating: "We are not providing data to any Chinese companies. They don't have access to any data… An open-source model does not have access to data. It doesn't work that way." Chesky said Airbnb is not a customer of Alibaba, meaning it uses Qwen open weights rather than Alibaba Cloud's hosted API, and that Airbnb primarily uses a variety of open-source models, including US open-source models. Airbnb's AI customer-service agent handles 70% of customer-service tickets (Source: bloomberg.com).
Chesky's statement was a CEO-level public articulation of the argument that open weights can be data-isolated, made in response to an April 29, 2026 probe by US House committees on China and Homeland Security into Airbnb and Anysphere/Cursor over their use of Chinese AI models. The committee letter cited national-security and data-security concerns about Airbnb's stated preference for "fast and cheap" Qwen in an October 2025 Chesky interview (Source: bloomberg.com).
In a Google I/O fireside chat the same week, Sundar Pichai voiced a similar view of open-weight provenance: "If it is open source with the right licenses, it should matter less where it came from… I worry less about, 'Are we adopting open source models from China?' and more, 'Are we doing enough in the US to make sure we are staying at the frontier?'" (Source: bloomberg.com).
Adoption of Chinese open-weight models extends beyond Airbnb. For May 2026, seven of the ten most-used models on OpenRouter, a multi-million-developer model-routing platform, were developed by Chinese companies, including DeepSeek, Moonshot AI, Tencent, and Qwen (Source: openrouter.ai; cited via bloomberg.com).
People
The Qwen3 report lists approximately 60 authors. Individual leadership attributions are not yet independently corroborated; specific names will be added when confirmed by additional sources. Contact details are not reproduced.
Sources and related pages
The primary technical source is Qwen3 Technical Report (arXiv 2505.09388, May 2025), the main technical reference for Qwen3.
Relationships
- related: DeepSeek, China and the US Are Running Different AI Races, Open-Source AI / Open-Weight Models, Distillation, Fast-Follow Problem, AI Sovereignty
- contradicts: the "Alibaba tilting closed" framing in Open-Source AI / Open-Weight Models and (Source: epochai.substack.com) — Qwen3's Apache 2.0 235B release is counter-evidence. The closed, API-only Qwen3.7-Max release (May 2026) aligns with the "tilting closed" framing at the flagship tier, even as the open Qwen3 family persists in lower tiers; Alibaba's posture is bifurcated (open mid-tier, closed flagship).