Downloads · 30 days
0
mpribadi/Docling-RapidOcr
Docling-RapidOcr is a image-text-to-text model from mpribadi. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for PaddleOCR. The card lists the license as apache-2.0.
This repository provides the missing PP-OCRv6 ONNX model files required by docling-serve) (v1.27.0 and later) for RapidOCR document extraction.
Downloads · 30 days
0
Access
Public
Updated Jul 29, 2026
Repo size
57.4 MB
Likes
0
Public
Click a slice to open those files.
.onnx57.4 MB · 100%
From the Hugging Face model README
This repository provides the missing PP-OCRv6 ONNX model files required by docling-serve (v1.27.0 and later) for RapidOCR document extraction.
When updating docling-serve to version 1.27.0+, the default model cache may only contain PP-OCRv4 weights, resulting in FileNotFoundError crashes when the engine attempts to load the v6 det_small and rec_small ONNX files. This repository consolidates these specific files to restore functionality for document extraction pipelines, such as those integrated with Open WebUI.
This repository contains two essential ONNX weights:
PP-OCRv6_det_small.onnxPP-OCRv6_rec_small.onnxThese are lightweight, highly optimized versions of the PaddleOCR v6 models, converted to ONNX format by the RapidOCR community for fast, CPU-friendly inference.
To resolve the FileNotFoundError in your docling-serve environment, you need to place these files into the internal cache directory expected by Docling.
The files must be placed in the following exact directory structure within the container or host cache (e.g., /opt/app-root/src/.cache/ or ~/.cache/):
.cache/docling/models/RapidOcr/onnx/PP-OCRv6/
├── det/
│ └── PP-OCRv6_det_small.onnx
└── rec/
└── PP-OCRv6_rec_small.onnx
docling-tools InstallationUse docling-tools cli command:
#targeting RapidOcr as output directory
docling-tools models download-hf-repo mpribadi/Docling-RapidOcr -o /home/runner/.cache/docling/models/RapidOcr/
If you are running the app directly on a host, download the files from this repository and move them to the corresponding path:
# Create the required directories
mkdir -p ~/.cache/docling/models/RapidOcr/onnx/PP-OCRv6/det
mkdir -p ~/.cache/docling/models/RapidOcr/onnx/PP-OCRv6/rec
# Move the downloaded models
mv PP-OCRv6_det_small.onnx ~/.cache/docling/models/RapidOcr/onnx/PP-OCRv6/det/
mv PP-OCRv6_rec_small.onnx ~/.cache/docling/models/RapidOcr/onnx/PP-OCRv6/rec/
This repository does not claim ownership of the original model architectures or weights. It serves solely as a convenient distribution mirror to solve a specific dependency issue in the Docling ecosystem.
Credit goes to the original authors and the communities hosting the specific ONNX exports:
PP-OCRv6_det_small.onnx): Sourced from the main branch of Hugging Face repository breezedeus/cnstd-ppocr-multi_PP-OCRv6_det_small.PP-OCRv6_rec_small.onnx): Sourced from the v3.9.1 branch of ModelScope repository RapidAI/RapidOCR.docling-serve podman container build for AMD Strix Halo (gfx1151) ➤ muslimpribadi/docling-serve-strix-halo