Downloads · 30 days
0
itextresearch/itext-EasyOCR-devanagari
itext-EasyOCR-devanagari is a machine learning model from itextresearch. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
These are machine learning models designed to detect and recognize text within images. They analyze visual input, identify regions containing text, and convert that text into a machine-readable format. We integrate th…
Downloads · 30 days
0
Access
Public
Updated Mar 23, 2026
Repo size
215 MB
Likes
0
Public
Click a slice to open those files.
.onnx215 MB · 100%
From the Hugging Face model README
These are machine learning models designed to detect and recognize text within images. They analyze visual input, identify regions containing text, and convert that text into a machine-readable format. We integrate these models into our iText PdfOCR ONNX engine to enable efficient and accurate extraction of textual content from scanned documents and image-based PDFs. For seamless integration, they are converted into ONNX format, ensuring compatibility with the iText PdfOCR engine.
<a href="https://itextpdf.com/products/pdf-ocr-text-recognition">Learn more about iText PdfOCR</a>
// detection
IDetectionPredictor easyDetectionPredictor = OnnxDetectionPredictor.easyOcr(PATH_TO_DETECTION_MODEL); // e.g. "itext-EasyOCR-craft_mlt_25k.onnx"
// recognition
IRecognitionPredictor recognition = OnnxRecognitionPredictor.easyOcr(PATH_TO_RECOGNITION_MODEL, EasyOcrMapper.LATIN_G2); // your recognition ONNX model + appropriate mapper
// combine into an engine
try(OnnxOcrEngine ocrEngine = new OnnxOcrEngine(detection, recognition);) {
/* Your OCR processing here */
}
<br>
<h4>OCR engine with orientation handling</h4>
<h5>java (but is also available in .NET)</h5>
IDetectionPredictor easyDetectionPredictor = OnnxDetectionPredictor.easyOcr(PATH_TO_DETECTION_MODEL);
IRecognitionPredictor recognition = OnnxRecognitionPredictor.easyOcr(PATH_TO_RECOGNITION_MODEL, EasyOcrMapper.LATIN_G2);
// orientation predictor (MobileNetV3 variant)
OnnxOrientationPredictor orientation = OnnxOrientationPredictor.mobileNetV3(PATH_TO_ORIENTATION_MODEL);
try(OnnxOcrEngine ocrEngine = new OnnxOcrEngine(detection, orientation, recognition)) {
/* Your OCR processing here */
}
<br>
<h4>The created ocrEngine can be passed to the standard iText pdfOCR API, e.g.:</h4>
OcrPdfCreator ocrPdfCreator = new OcrPdfCreator(ocrEngine);
ocrPdfCreator.createPdfFile(Collections.singletonList(new File(imagePath)), new File(outPdfFile))