Ad
Skip to content

Alibaba releases Qwen3 compact open source multimodal models

Alibaba's Qwen group has released two new small-scale multimodal models, Qwen3-VL-30B-A3B-Instruct and Qwen3-VL-30B-A3B-Thinking, each with 3 billion active parameters. According to Qwen, both versions are competitive with GPT-5-Mini and Claude 4 Sonnet, and in some benchmarks show stronger performance in math, image recognition, text recognition, video processing, and agent control.

The lineup includes an FP8 version for faster inference, and an FP8 variant of the Qwen3-VL-235B-A22B model. The models are available on HuggingFace, ModelScope, and GitHub, or via an Alibaba Cloud API. There is also a web chat interface for direct use.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

Read on for the full picture.
Subscribe for hype-free coverage.

  • Full access to every article on THE DECODER
  • No ads
  • Join the comments and community discussions
  • A weekly AI news recap via mail
  • 6x/year: "AI Radar" — deep dives on the AI topics that matter most
  • Daily AI news, always up to date
  • Our full ten-year archive
  • Covered by a team with 10+ years in AI
Subscribe to The Decoder