Category: Alibaba models

This category contains 11 pages.

  • Qwen (model family) · Alibaba's open-weight model line; by 2025 the most-downloaded and most-fine-tuned base family in the open ecosystem.
  • Qwen 1 · Alibaba Cloud's first generation of Qwen large language models (2023), released as open-weight checkpoints from 1.8B to 72B parameters.
  • Qwen2 · Alibaba's June 2024 open-weight model series spanning 0.5B to 72B parameters, including one mixture-of-experts variant.
  • Qwen2.5 · Alibaba's September 2024 open-weight model series (0.5B to 72B), trained on a developer-reported 18 trillion tokens and widely used as a base for fine-tunes and reasoning distillations.
  • Qwen2.5-Max · Alibaba's January 2025 proprietary large-scale mixture-of-experts flagship, positioned against DeepSeek-V3 and Western frontier models.
  • Qwen2.5-VL · Alibaba's January 2025 open-weight vision-language model series (3B to 72B) with document parsing, object grounding, long-video understanding, and GUI-agent capabilities.
  • Qwen3 · Alibaba's April 2025 open-weight model family introducing hybrid thinking and non-thinking modes across dense and mixture-of-experts variants.
  • Qwen3-Coder · Alibaba's July 2025 open-weight agentic coding model, a 480B-parameter mixture-of-experts specialist derived from the Qwen3 family.
  • Qwen3-Coder-Next · Alibaba Qwen team's February 2026 open-weight coding model, an ultra-sparse 80B hybrid-attention MoE tuned for coding agents and local development.
  • Qwen3-Max · Alibaba's September 2025 proprietary flagship of the Qwen3 generation, developer-reported to exceed one trillion parameters.
  • QwQ-32B · Alibaba Qwen team's open-weight 32B reasoning model, previewed in November 2024 and released in March 2025 with developer-reported parity to much larger reasoners.