PaddleOCR-VL
A compact document parsing model from the popular PaddleOCR toolkit. Per the model card it supports 109 languages, including Russian; version 1.6 leads the OmniDocBench benchmark.
- Developer
- Baidu (PaddlePaddle), China
- First release
- Oct 2025
- Latest release
- May 2026
- Sizes
- 0.9B
- License
- Commercial use allowedApache 2.0
- Russian
- Supported
- Ready-made builds
- GGUF, MLX (Apple)
- Running
- On your own serverAlso runs without a GPU
- Industries
- Documents and accounting, Finance, Legal
What it does
- Recognising invoices, contracts and delivery notes, including in Russian
- Recognising tables, formulas and stamps
- Converting scans to Markdown and JSON
- Running on an ordinary PC without a powerful GPU
Where it is used
Hardware requirements
Versions
- PaddleOCR-VL-1.6
- PaddleOCR-VL-1.5
- PaddleOCR-VL
How to run it
I can set this up end to end: pick the model size, deploy it on your server and connect it to your systems. Quantization compresses a model so it takes less video memory and runs on more modest hardware. Answers change slightly, so quality is checked on your own examples.
Frequently asked questions
Can PaddleOCR-VL be used in a commercial project?
Yes. License: Apache 2.0. It allows commercial use, but it is still worth having a lawyer review the license before launch.
What hardware does PaddleOCR-VL need?
At minimum: Laptop or regular PC, up to 8 GB of VRAM — smaller versions. Some versions also run on an ordinary CPU, without a GPU. You can calculate the exact VRAM for your model size and context in the hardware calculator.
Does PaddleOCR-VL support Russian?
Yes, Russian is listed on the model card.
Where can I download PaddleOCR-VL and what does it cost?
The PaddleOCR-VL weights are open and free to download. You only pay for the hardware it runs on and for the setup. Source links are at the bottom of this page.
How I deploy it for clients
- SelectionI pick the model size for your task and hardware and test it on your examples.
- DeploymentI deploy it on your server or in a closed network and provide an API.
- Fine-tuningI fine-tune it on your data (LoRA) or connect a knowledge base — whichever is cheaper for the task.
- IntegrationI connect it to your CRM, ERP, bot, website or team chat and set up monitoring.
Comparisons
Similar models
A lightweight OCR model from Zhipu for document parsing. The model card lists Russian among supported languages; built for high load and low-end hardware.
DetailsDocuments and OCRMinerUShanghai AI Laboratory (OpenDataLab) · ChinaCommercial use allowedA popular open tool for converting PDFs to Markdown with its own small model. MinerU2.5-Pro was improved through data alone, without growing in size. Languages on the card: Chinese and English.
DetailsDocuments and OCRQianfan-OCRBaidu (Qianfan) · ChinaCommercial use allowedA Baidu model that not only recognises a document but also answers questions about it. Per the model card it supports 192 languages, including Cyrillic.
DetailsSource: huggingface.co/PaddlePaddle/PaddleOCR-VL-1.6. Data checked against the model card on 22 Sep 2026. Have a lawyer review the license before commercial launch.


