Model reference · open weights
UVDoc is an open-weight language model from PaddlePaddle, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.
About
UVDoc Introduction The main purpose of text image correction is to carry out geometric transformation on the image to correct the document distortion, inclination, perspective deformation and other problems in the image, so that the subsequent text recognition can be more accurate. Note: Test data set: docunet benchmark data set. Quick Start Installation 1. PaddlePaddle Please refer to the following commands to install PaddlePaddle using pip: For details about PaddlePaddle installation, please refer to the PaddlePaddle official website. 2. PaddleOCR Install the latest version of the PaddleOCR inference package from PyPI: Model Usage You can quickly experience the functionality with a single command: You can also integrate the model inference of the TextImageUnwarping module into your project. Before running the following code, please download the sample image to your local machine. After running, the obtained result is as follows: The visualized image is as follows: For details about usage command and descriptions of parameters, please refer to the Document. Pipeline Usage The ability of a single model is limited. But the pipeline consists of several models can provide more capacity to resolve difficult problems in real-world scenarios. PP-StructureV3 Layout analysis is a technique used to extract structured information from document images. PP-StructureV3 includes the following six modules: Layout Detection Module General OCR Sub-pipeline Document Image Preprocessing Sub-pipeline (Optional) Table Recognition Sub-pipeline (Optional) Seal Recognition Sub-pipeline (Optional) Formula Recognition Sub-pipeline (Optional) You can quickly experience the PP-StructureV3 pipeline with a single command. You can experience the inference of the pipeline with just a few lines of code. Taking the PP-StructureV3 pipeline as an example: For details about usage command and descriptions of parameters, please refer to the Document. Links PaddleOCR Repo PaddleOCR Documentation
Summarised from the published model card. Read the full card on the HuggingFace links below.
Specifications
| Maker | PaddlePaddle |
|---|---|
| Type | Language models |
| Variants | 1 |
| Runs with | PaddleOCR |
| Released | 2025-06-06 |
| Popularity | 426k downloads / month |
| Likes | 12 |
| Licence | Open weights |
How it works
Variants
Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.
| Variant | Params | Precision | VRAM | Fits 16 GB | Weights |
|---|---|---|---|---|---|
| UVDoc | — | BF16 | — | — | Weights ↗ |
Using it via the API
Once AxForge deploys uvdoc for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (uvdoc below is illustrative; you get the exact model name on deployment.)
$ curl -sS https://api.axforge.ai/v1/chat/completions \
-H "Authorization: Bearer $AXFORGE_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"uvdoc","messages":[{"role":"user","content":"Hello"}]}'
Details
Languages
Tags
Licence
Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗