Models & Providers Entry

Qwen, explained

Reviewed October 2026

TL;DR: Qwen is Alibaba's model family, released mostly with open weights across an unusually wide range of sizes and modalities. It is one of the most-downloaded and most-fine-tuned base families in the open ecosystem, with particular strength in multilingual and coding work. You can download it, call Alibaba Cloud's API, or use third-party hosts.

What it is and how you use it

Qwen is built by Alibaba's Qwen team, inside the Chinese e-commerce and cloud group. Where some labs ship one flagship at a time, Qwen's strategy is breadth: each generation arrives as a lineup spanning tiny models for edge devices up to large mixture-of-experts systems, plus specialized branches for coding, vision, audio, and math. Most of the lineup ships with open weights, commonly under the permissive Apache 2.0 license - though the newest top-tier open releases carry Qwen's own licenses with revenue-scale conditions attached, so the license is worth checking per checkpoint rather than assumed. The largest tiers were once held back as API-only; Qwen now open-releases a flagship-class checkpoint alongside the smaller sizes, under its own license rather than Apache 2.0, while the hosted top tier stays closed.

That breadth made Qwen infrastructure for other people's models. Because capable, permissively licensed checkpoints exist at nearly every size, Qwen is one of the most popular bases on Hugging Face for fine-tuning - many well-known community and research models are Qwen underneath. The family is also known for strong multilingual coverage, including but not limited to Chinese, and for coding variants that hold their own in developer tooling.

Access mirrors the other open families. Download the weights and run small variants locally through Ollama or serve larger ones on your own GPUs; call Alibaba's hosted API - QwenCloud internationally, the Qianwen AI Platform in mainland China, both a 2026 rebrand of the service its own docs and SDKs still call DashScope or Model Studio; or use Western inference providers that host the open checkpoints on their own hardware. Alibaba runs the platform in international regions as well as mainland ones, so the first-party path is open to more teams than a China-only endpoint would be - and where it is not, the open license is what makes self-hosting or a third-party host available at all.

Where it typically fits: multilingual products, fine-tuned domain models that need a permissive base, and edge or on-device deployments that depend on the family's smallest tiers. Its breadth is the draw - whatever size and modality your constraint dictates, there is usually a Qwen checkpoint shaped like it - while the usual open-model trade-off applies: you or your provider carry the serving burden the closed labs would otherwise hide.

Where it sits in the AI stack

Qwen sits at the model layer as an open family: pick a size and branch, then pick who serves it - your own hardware, Alibaba Cloud, or a third party:

Key tools and implementations

  • Open-weight downloads

    Checkpoints at nearly every size on Hugging Face - most under Apache 2.0, the largest under Qwen's own licenses.

  • QwenCloud

    Alibaba's hosted platform, formerly Model Studio, serves the full Qwen lineup and some third-party open models, for teams that want them without running them.

  • Third-party providers

    Western inference providers hosting the open checkpoints for teams that avoid China-based endpoints.

  • Local runtimes

    Ollama and llama.cpp run the small and mid-size variants on ordinary desktops and laptops.