Downloads · 30 days
8
2% of all-time downloads
eduardofarina/MultimodalXray
MultimodalXray is a image-text-to-text model from eduardofarina. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
This model is trained on a sample of CheXpert Dataset using only frontal views. The model combines a ViT and GPT2 to generate draft radiology reports.
Downloads · 30 days
8
2% of all-time downloads
All-time downloads
402
Public
Repo size
21.2 GB
Likes
1
Public
Click a slice to open those files.
.pt13.4 GB · 63%
From the Hugging Face model README
This model is trained on a sample of CheXpert Dataset using only frontal views. The model combines a ViT and GPT2 to generate draft radiology reports.