Skip to content

priyank-m

m_OCR

priyank-m/m_OCR

m_OCR is a image-text-to-text model from priyank-m. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.

Multilingual OCR (mOCR) is a VisionEncoderDecoder model based on the concept of TrOCR for English and Chinese document text-recognition. It uses a pre-trained Vision encoder and a pre-trained Language model as decoder.

Downloads · 30 days

93

4% of all-time downloads

All-time downloads

2.3K

Public

Repo size

36.6 GB

Likes

13

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin2.4 GB · 100%

At a glance

Task
Image-Text-to-Text
Library
transformers
Model type
vision-encoder-decoder
Access
Public
Created
Nov 24, 2022
Updated
Jan 9, 2023
SHA
01f884da

Try a prompt

Task
Image-Text-to-Text
Library
transformers
Type
vision-encoder-decoder
Languages
en, zh, multilingual
Created
Nov 24, 2022
Updated
Jan 9, 2023