All models

Qwen3-VL-235B-A22B-Instruct

byQwenQwen· 23 Sep 2025
General purposeMultimodal

The Qwen VL series finally gets its long-awaited and anticipated update with small (4B, 8B) dense and larger (30B-A3B, 235B-A22B) MoEs in both instruct and reasoning versions. We want to shine a special spotlight on the 8B variants: Their text benchmarks have also improved across the board compared to the initial 8B release - reinforcing our point on the challenge of hybrid reasoning. As the 8B versions did not get a 2507 refresh, these versions should be a no-brainer update and drop-in replacement if you were using Qwen3 8B (or are still using Llama3.1 8B).

Specs
Params235B, 22B active
LicenseApache-2.0
Capability · Artificial Analysis
AA Index
14.4
Months Behind Frontier
12.4 mo
o1-preview
12 Sep 2024
Adoption · Hugging Face
RAM @ 30d
Relative Adoption Metric: 0.46×. Below benchmark. Measured at 30 days.
Hugging Face Downloads
1.3M
last 30d
9.1M
all time
HF Likes
415

Relative Adoption Metric contextualizes downloads against the model's size bucket.

Inference · OpenRouter
Tokens/Day
below top 50 this week
Peak Tokens/Day
12.8B
10 Feb 2026
Peak Rank
#18
12 Oct 2025

OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Qwen3-VL-235B-A22B-Instruct fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.

Related Models