Skip to content

pandafm

donut-es

pandafm/donut-es

donut-es is a image-text-to-text model from pandafm. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.

This is a fine-tuned version of the Donut architecture, specifically tailored for parsing retail receipts. Donut is a transformer-based model designed for document understanding, and it performs OCR-free parsing by di…

Downloads · 30 days

27

1% of all-time downloads

All-time downloads

2.3K

Public

Parameters

201M

20.2 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors806 MB · 99%

Parameter types

How the weights are stored.

F32201M · 100%

Base models

Task
Image-Text-to-Text
Library
transformers
Type
vision-encoder-decoder
License
mit
Languages
es
Created
May 7, 2024
Updated
Oct 13, 2024
donut-es — AI Model — AIMarketly