Downloads · 30 days
1
7% of all-time downloads
qhfmshal/TRPaliGemma
TRPaliGemma is a image-text-to-text model from qhfmshal. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.
This model is fine-tuned PaliGemma model for the Table recognition task.
Downloads · 30 days
1
7% of all-time downloads
All-time downloads
15
Public
Repo size
610 MB
Likes
0
Public
Click a slice to open those files.
.safetensors45.3 MB · 67%
From the Hugging Face model README
This model is fine-tuned PaliGemma model for the Table recognition task.
<!-- Provide a quick summary of what the model is/does. -->Table recognition is a branch of Document AI. In the existing Table recognition, the structure of the table and the OCR results were calculated and combined, respectively. For this reason, unnecessary predictions are sometimes made in the process of parsing the table.(ex. bbox) Using VLM, the structure and text of the table will be predicted at the same time, eliminating unnecessary predictions and integrating the two tasks into one.
<!-- Provide a longer summary of what this model is. -->This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated.
This model can convert a tabular images into HTML.
It can be used in document automation systems using Document AI.
This is a fine-tuned model with only the tabular images that exist within the PDF, so you won't get good performance in the tabular images in the wild.
This model simply converts table images into HTML. To gain additional analysis or knowledge, you need to learn an NLP model for analysis using HTML or fine-tune the new PaliGemma model by constructing new data.
inference : https://www.kaggle.com/code/mldlchoidh/tr-inference
Pubtables1-1M