Skip to content

TIGER-Lab

VLM2Vec-Qwen2VL-7B

TIGER-Lab/VLM2Vec-Qwen2VL-7B

VLM2Vec-Qwen2VL-7B is a image-text-to-text model from TIGER-Lab. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.

A new checkpoint trained using Qwen/Qwen2-VL-7B-Instruct with an enhanced training setup (LoRA tuning, batch size of 2048, maximum sub-dataset size of 100k). This model has shown significantly improved performance on…

Downloads · 30 days

277

1% of all-time downloads

All-time downloads

40.8K

Public

Repo size

33.6 GB

Likes

12

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin83.7 MB · 82%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
apache-2.0
Model type
qwen2_vl
Access
Public
Created
Feb 24, 2025
Updated
May 3, 2025
SHA
5aec77d4

Try a prompt

Base models

Task
Image-Text-to-Text
Library
transformers
Type
qwen2_vl
License
apache-2.0
Languages
en
Created
Feb 24, 2025
Updated
May 3, 2025