Qwen-VL
One of the strongest open vision models: reads documents, tables, charts and video, and works with user interfaces. Since Qwen3.5, vision is built directly into the main Qwen model.
- Developer
- Alibaba (Qwen team), China
- First release
- Aug 2023
- Latest release
- Oct 2025
- Sizes
- 2B – 235B-A22B
- License
- Commercial use allowedQwen3-VL: Apache 2.0; older Qwen-VL and some Qwen2/2.5-VL models (3B, 72B) have their own Qwen licenses
- Russian
- Not stated
- Ready-made builds
- GGUF
- Running
- Available in OllamaAlso runs without a GPU
- Industries
- Documents and accounting, Retail and marketplaces, Customer support, Manufacturing and logistics
What it does
- Extracting data from scanned invoices and delivery notes
- Analysing photos of products and shelves
- Analysing video and camera footage
- An agent that operates an interface from screenshots
Where it is used
Hardware requirements
Versions
- Qwen3-VL 2B – 32B
- Qwen3-VL 235B-A22B и 30B-A3B
- Qwen2.5-VL-32B
- Qwen2.5-VL
- Qwen2-VL
- Qwen-VL
How to run it
I can set this up end to end: pick the model size, deploy it on your server and connect it to your systems. Quantization compresses a model so it takes less video memory and runs on more modest hardware. Answers change slightly, so quality is checked on your own examples.
Frequently asked questions
Can Qwen-VL be used in a commercial project?
Yes. License: Qwen3-VL: Apache 2.0; older Qwen-VL and some Qwen2/2.5-VL models (3B, 72B) have their own Qwen licenses. It allows commercial use, but it is still worth having a lawyer review the license before launch.
What hardware does Qwen-VL need?
At minimum: Laptop or regular PC, up to 8 GB of VRAM — smaller versions. Some versions also run on an ordinary CPU, without a GPU. You can calculate the exact VRAM for your model size and context in the hardware calculator.
Does Qwen-VL support Russian?
The model card does not list languages, so Russian support cannot be promised — it has to be tested on your own examples.
Where can I download Qwen-VL and what does it cost?
The Qwen-VL weights are open and free to download. You only pay for the hardware it runs on and for the setup. Source links are at the bottom of this page.
How I deploy it for clients
- SelectionI pick the model size for your task and hardware and test it on your examples.
- DeploymentI deploy it on your server or in a closed network and provide an API.
- Fine-tuningI fine-tune it on your data (LoRA) or connect a knowledge base — whichever is cheaper for the task.
- IntegrationI connect it to your CRM, ERP, bot, website or team chat and set up monitoring.
Comparisons
Similar models
A family of language models with strong Russian language support, from small versions for a laptop to a flagship on par with commercial APIs.
DetailsImage + textInternVLShanghai AI Laboratory (OpenGVLab) · ChinaCommercial use allowedA large family of Chinese vision models sized from 1B to 241B. InternVL-U (4B) combines image understanding, generation and editing.
DetailsImage + textMiniCPM-VOpenBMB (ModelBest and Tsinghua University) · ChinaCommercial use allowedCompact vision models that run even on a phone or laptop. Good at reading text in photos and understanding video; version 4.6 is only 1.3B.
DetailsSource: huggingface.co/Qwen/Qwen3-VL-8B-Instruct. Data checked against the model card on 22 Sep 2026. Have a lawyer review the license before commercial launch.


