Downloads · 30 days
0
baudm/abinet-lv
abinet-lv is a image-to-text model from baudm. Use it when you need a caption or text from an image. The card lists the license as apache-2.0.
ABINet model pre-trained on various real STR datasets at image size 128x32.
Downloads · 30 days
0
Access
Public
Updated Aug 28, 2022
Repo size
148 MB
Likes
0
Public
Click a slice to open those files.
.bin148 MB · 100%
From the Hugging Face model README
ABINet model pre-trained on various real STR datasets at image size 128x32.
Disclaimer: this model card was not written by the original authors.
TODO
You can use the model for STR on images containing Latin characters (62 case-sensitive alphanumeric + 32 punctuation marks).
TODO
@InProceedings{Fang_2021_CVPR,
author = {Fang, Shancheng and Xie, Hongtao and Wang, Yuxin and Mao, Zhendong and Zhang, Yongdong},
title = {Read Like Humans: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Recognition},
booktitle = {Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)},
month = {6},
year = {2021},
pages = {7098-7107}
}