Downloads · 30 days
0
keras-io/image-captioning
image-captioning is a image-to-text model from keras-io. Use it when you need a caption or text from an image. It is set up for generic. The card lists the license as cc0-1.0.
This repo contains the models and the notebook on Image captioning with visual attention.
Downloads · 30 days
0
Access
Public
Updated Jan 13, 2022
Repo size
63.7 MB
Likes
8
Public
Click a slice to open those files.
.h526.5 MB · 44%
From the Hugging Face model README
This repo contains the models and the notebook on Image captioning with visual attention.
Full credits to TensorFlow Team
This notebook implements TensorFlow Keras implementation on Image captioning with visual attention.
Given an image like the example below, your goal is to generate a caption such as "a surfer riding on a wave".
To accomplish this, you'll use an attention-based model, which enables us to see what parts of the image the model focuses on as it generates a caption.
The model architecture is similar to Show, Attend and Tell: Neural Image Caption Generation with Visual Attention.
This notebook is an end-to-end example. When you run the notebook, it downloads the MS-COCO dataset, preprocesses and caches a subset of images using Inception V3, trains an encoder-decoder model, and generates captions on new images using the trained model.