Skip to content

inference-net

ClipTagger-12b

inference-net/ClipTagger-12b

ClipTagger-12b is a image-text-to-text model from inference-net. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.

Downloads · 30 days

60

1% of all-time downloads

All-time downloads

6.6K

Public

Parameters

12.2B

13.7 GB on disk

Likes

59

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors13.6 GB · 100%

Parameter types

How the weights are stored.

F8_E4M310.8B · 88%

Try a prompt

Base models

Task
Image-Text-to-Text
Type
gemma3
License
apache-2.0
Languages
en
Created
Aug 13, 2025
Updated
Aug 14, 2025
ClipTagger-12b — AI Model — AIMarketly