Skip to content

ziiiteng

CLIP-IN

ziiiteng/CLIP-IN

CLIP-IN is a zero-shot image classification model from ziiiteng. Use it for the zero-shot image classification task on the model card, and read the license before you ship it in a product. The card lists the license as mit.

Despite the success of Vision-Language Models (VLMs) like CLIP in aligning vision and language, their proficiency in detailed, fine-grained visual comprehension remains a key challenge. We present CLIP-IN, a novel fra…

Downloads · 30 days

0

Access

Public

Updated Oct 14, 2025

Repo size

32.4 GB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pt32.4 GB · 100%

At a glance

Task
Zero-Shot Image Classification
License
mit
Access
Public
Created
Oct 13, 2025
Updated
Oct 14, 2025
SHA
88600062

Base models

Task
Zero-Shot Image Classification
License
mit
Languages
en
Created
Oct 13, 2025
Updated
Oct 14, 2025