Downloads · 30 days
7
1% of all-time downloads
ayoubkirouane/git-base-One-Piece
git-base-One-Piece is a image-to-text model from ayoubkirouane. Use it when you need a caption or text from an image. It is set up for transformers. The card lists the license as mit.
+ Model Name: Git-base-One-Piece + Base Model: Microsoft's "git-base" model + Model Type: Generative Image-to-Text (GIT) + Fine-Tuned On: 'One-Piece-anime-captions' dataset + Fine-Tuning Purpose: To generate text capt…
Downloads · 30 days
7
1% of all-time downloads
All-time downloads
548
Public
Repo size
1.4 GB
Likes
0
Public
Click a slice to open those files.
.bin707 MB · 100%
From the Hugging Face model README
Git-base-One-Piece is a fine-tuned variant of Microsoft's git-base model, specifically trained for the task of generating descriptive text captions for images from the One-Piece-anime-captions dataset.
The dataset consists of 856 {image: caption} pairs, providing a substantial and diverse training corpus for the model.
The model is conditioned on both CLIP image tokens and text tokens and employs a teacher forcing training approach. It predicts the next text token while considering the context provided by the image and previous text tokens.

# Use a pipeline as a high-level helper
from transformers import pipeline
pipe = pipeline("image-to-text", model="ayoubkirouane/git-base-One-Piece")
or
# Load model directly
from transformers import AutoProcessor, AutoModelForCausalLM
processor = AutoProcessor.from_pretrained("ayoubkirouane/git-base-One-Piece")
model = AutoModelForCausalLM.from_pretrained("ayoubkirouane/git-base-One-Piece")