Skip to content

ayushadarsh7

gemma_3_4b_only_vision

ayushadarsh7/gemma_3_4b_only_vision

gemma_3_4b_only_vision is a image-text-to-text model from ayushadarsh7. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.

This model is a fine-tuned version of google/gemma-3-4b-it. It has been trained using TRL.

Downloads · 30 days

14

33% of all-time downloads

All-time downloads

43

Public

Parameters

4.3B

17.2 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors17.2 GB · 100%

At a glance

Task
Image-Text-to-Text
Library
transformers
Model type
gemma3
Access
Public
Created
Nov 25, 2025
Updated
Nov 26, 2025
SHA
bea00f04

Try a prompt

Base models

Task
Image-Text-to-Text
Library
transformers
Type
gemma3
Created
Nov 25, 2025
Updated
Nov 26, 2025