Downloads · 30 days
15
1% of all-time downloads
Brian314/TexTeller
TexTeller is a image-to-text model from Brian314. Use it when you need a caption or text from an image. The card lists the license as apache-2.0.
Downloads · 30 days
15
1% of all-time downloads
All-time downloads
1.6K
Public
Parameters
298M
15.1 GB on disk
Likes
1
Public
Click a slice to open those files.
.onnx3 GB · 56%
From the Hugging Face model README
There are more test images here and a horizontal comparison of recognition models from different companies.
TexTeller is a ViT-based model designed for end-to-end formula recognition. It can recognize formulas in natural images and convert them into LaTeX-style formulas.
TexTeller is trained on a larger dataset of image-formula pairs (a 550K dataset available here), exhibits superior generalization ability and higher accuracy compared to LaTeX-OCR, which uses approximately 100K data points. This larger dataset enables TexTeller to cover most usage scenarios more effectively.
For more details, please refer to the 𝐓𝐞𝐱𝐓𝐞𝐥𝐥𝐞𝐫 GitHub repository.