Context
131K tokens (~66 books)
Input $/1M
$0.10
Output $/1M
$0.42
Type
multimodal
License
Open Source
Benchmarks
0 tested
Data as of
About
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
No benchmark data available yet.
Links
Research
Documentation
Community
Source Code
BenchGecko API
qwen3-vl-32b-instruct
Specifications
- Typemultimodal
- Context131K tokens (~66 books)
- ReleasedOct 2025
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.001
Available On
Share & Export
Frequently Asked Questions
Qwen3 VL 32B Instruct is an open-source multimodal AI model by Alibaba Qwen, released in October 2025. Context window: 131K tokens.
Related Models
Qwen3 Coder Next FP8 · AlibabaQwen2.5 Coder 32B Instruct AWQ · AlibabaSpeaker Diarization Community 1 · PyannoteGemma 3 12B (free) · Google DeepMindQwen3 Next 80B A3B Instruct (free) · Alibaba QwenBenchmarks
Key facts · as of 2026-10-05
- Qwen3 VL 32B Instruct by Alibaba Qwen. Not enough public benchmark scores to rank yet.
- List price $0.10 input · $0.42 output per 1M tokens (as of 2026-10-05).
- Sold by 1 provider (as of 2026-10-05): Alibaba $0.10 in / $0.42 out.
How to cite · data as of 2026-10-05
Qwen3 VL 32B Instruct · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/qwen3-vl-32b-instruct
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP