Downloads · 30 days
12.2K
3% of all-time downloads
microsoft/layoutlm-large-uncased
layoutlm-large-uncased is a machine learning model from microsoft. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
Multimodal (text + layout/format + image) pre-training for document AI
Downloads · 30 days
12.2K
3% of all-time downloads
All-time downloads
402K
Public
Repo size
4.1 GB
Likes
10
Public
Click a slice to open those files.
.bin1.4 GB · 50%
From the Hugging Face model README
Multimodal (text + layout/format + image) pre-training for document AI
Microsoft Document AI | GitHub
LayoutLM is a simple but effective pre-training method of text and layout for document image understanding and information extraction tasks, such as form understanding and receipt understanding. LayoutLM archives the SOTA results on multiple datasets. For more details, please refer to our paper:
LayoutLM: Pre-training of Text and Layout for Document Image Understanding Yiheng Xu, Minghao Li, Lei Cui, Shaohan Huang, Furu Wei, Ming Zhou, KDD 2020
We pre-train LayoutLM on IIT-CDIP Test Collection 1.0* dataset with two settings.
If you find LayoutLM useful in your research, please cite the following paper:
@misc{xu2019layoutlm,
title={LayoutLM: Pre-training of Text and Layout for Document Image Understanding},
author={Yiheng Xu and Minghao Li and Lei Cui and Shaohan Huang and Furu Wei and Ming Zhou},
year={2019},
eprint={1912.13318},
archivePrefix={arXiv},
primaryClass={cs.CL}
}