Skip to content

sovitrath

Phi-3.5-vision-instruct

sovitrath/Phi-3.5-vision-instruct

Phi-3.5-vision-instruct is a image-text-to-text model from sovitrath. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.

This is an updated version of Phi 3.5 Vision Instruct which is supported by the latest version of Transformers. Checked with Transformers version 4.51.3. This line needed changing in modelingphi3v.py.

Downloads · 30 days

52

11% of all-time downloads

All-time downloads

489

Public

Parameters

4.1B

8.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors8.3 GB · 100%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
mit
Model type
phi3_v
Access
Public
Created
May 11, 2025
Updated
May 11, 2025
SHA
c9d92840
Task
Image-Text-to-Text
Library
transformers
Type
phi3_v
License
mit
Languages
multilingual
Created
May 11, 2025
Updated
May 11, 2025
Phi-3.5-vision-instruct — AI Model — AIMarketly