All models

Qwen3.5-397B-A17B

byQwenQwen· 16 Feb 2026
General purposeMultimodal

The long-awaited update to Qwen is finally here. It comes in various sizes from 0.8B to 27B (dense) and 35B-A3B to 397B-A17B (MoE), some of them even with base models. All of them are multi-modal, use reasoning by default and are based on the Qwen-Next architecture with GDN layers.

We tested these models over the last few days, and they are a clear upgrade over the previous version: There are a lot of substantial improvements across the board, making them perfect workhorses for a wide range of tasks.
Their style and instruction-following have improved, and the models are even better at multilingual tasks, covering more languages.

However, at least the small models (still) tend to overthink. You can turn off reasoning by disabling it in the chat template.

Specs
Params397B
LicenseApache-2.0
Capability · Artificial Analysis
AA Index
34.3
Months Behind Frontier
6.4 mo
claude-4-1-opus-thinking
05 Aug 2025
Adoption · Hugging Face
RAM @ 30d
Relative Adoption Metric: 3.47×. Strong. Measured at 30 days.
Hugging Face Downloads
243.1K
last 30d
5.2M
all time
HF Likes
1.6K

Relative Adoption Metric contextualizes downloads against the model's size bucket.

Inference · OpenRouter
Tokens/Day
below top 50 this week
Peak Tokens/Day
42.7B
12 May 2026
Peak Rank
#24
04 Mar 2026

OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Qwen3.5-397B-A17B fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.

Related Models