Qwen
Qwen: Qwen3 VL 32B Instruct
Qwen
Input / M
Routed
Output / M
Context
Max Output
33K
Usage
Unranked
Quality
Unscored
Providers
1
Best Uptime 30m
100.00%
Latency (p50)
1192 ms
Throughput
43.0 tok/s
Released
October 23, 2025
Modalities
text, image
About
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text.
Capabilities
Vision
Tools
Reasoning
Streaming
JSON Mode
Host offerings
Per-host pricing, uptime, latency
No host data yet — will populate on next sync.
Price History
0 recorded changes
No price changes recorded yet.
Activity Timeline
- No events yet.
Versus
Nearest peers by price, context, and capabilities. One click to compare.
No comparable models found.
Show this WM Score
Free live badge for your README, docs or site
Live badge — it always shows the current WM Score. Free to use anywhere, no key needed.
Markdown
[](https://whatmodel.app/models/c66728cc-6fe0-4336-969d-467db017b140)HTML
<a href="https://whatmodel.app/models/c66728cc-6fe0-4336-969d-467db017b140"><img src="https://whatmodel.app/api/public/embed/score/c66728cc-6fe0-4336-969d-467db017b140" alt="Qwen: Qwen3 VL 32B Instruct WM Score" height="26" /></a>Image URL
https://whatmodel.app/api/public/embed/score/c66728cc-6fe0-4336-969d-467db017b140First seen 43d ago · Last verified 4m ago
