Skip to content

katanaml

donut-demo

katanaml/donut-demo

donut-demo is a image-to-text model from katanaml. Use it when you need a caption or text from an image. It is set up for transformers. The card lists the license as mit.

Donut model finetuned with CORD dataset. Mean accuracy: 0.901198895445646

Downloads · 30 days

17

1% of all-time downloads

All-time downloads

1.7K

Public

Repo size

25.1 GB

Likes

3

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin809 MB · 99%

At a glance

Task
Image-to-Text
Library
transformers
License
mit
Model type
vision-encoder-decoder
Access
Public
Created
Jan 18, 2023
Updated
Jan 19, 2023
SHA
7213e17d
Task
Image-to-Text
Library
transformers
Type
vision-encoder-decoder
License
mit
Created
Jan 18, 2023
Updated
Jan 19, 2023