Downloads · 30 days
4
4% of all-time downloads
MWirelabs/mizo-ocr
mizo-ocr is a image-to-text model from MWirelabs. Use it when you need a caption or text from an image. The card lists the license as cc-by-4.0.
The first OCR model for the Mizo language, developed by MWire Labs.
Downloads · 30 days
4
4% of all-time downloads
All-time downloads
94
Public
Parameters
334M
1.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.3 GB · 100%
From the Hugging Face model README
The first OCR model for the Mizo language, developed by MWire Labs.
MizoOCR is a fine-tuned TrOCR model for recognizing printed Mizo text, including its unique diacritical characters (â, ê, î, ô, û). It is built on microsoft/trocr-base-printed and trained on 70,000 deduplicated mix of curated + synthetic image-text pairs drawn from a 200k dataset generated by MWire Labs.
| Split | Character Accuracy |
|---|---|
| Validation | 89.61% |
| Test | 90.68% |
from transformers import TrOCRProcessor, VisionEncoderDecoderModel
from PIL import Image
processor = TrOCRProcessor.from_pretrained("MWirelabs/mizo-ocr")
model = VisionEncoderDecoderModel.from_pretrained("MWirelabs/mizo-ocr")
image = Image.open("mizo_text.jpg").convert("RGB")
pixel_values = processor(image, return_tensors="pt").pixel_values
generated = model.generate(pixel_values)
text = processor.tokenizer.decode(generated[0], skip_special_tokens=True)
print(text)
If you use this model, please cite:
@misc{mwirelabs2026mizoocr,
title={MizoOCR: First OCR Model for the Mizo Language},
author={MWire Labs},
year={2026},
publisher={Hugging Face},
url={https://huggingface.co/MWirelabs/mizo-ocr}
}
MWire Labs is an AI research organization based in Shillong, Meghalaya, India, specializing in language technology for Northeast India's indigenous languages.