Skip to content

ServiceNow

Llama-3.2-11B-Vision-Instruct-StarFlow

ServiceNow/Llama-3.2-11B-Vision-Instruct-StarFlow

Llama-3.2-11B-Vision-Instruct-StarFlow is a image-text-to-text model from ServiceNow. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as llama3.2.

Llama-3.2-11B-Vision-Instruct-StarFlow is a vision-language model finetuned for structured workflow generation from sketch images. It translates hand-drawn or computer-generated workflow diagrams into structured JSON…

Downloads · 30 days

14

4% of all-time downloads

All-time downloads

365

Public

Parameters

10.7B

21.4 GB on disk

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors21.3 GB · 100%

At a glance

Task
Image-Text-to-Text
Library
transformers
License
llama3.2
Model type
mllama
Access
Public
Created
May 26, 2025
Updated
Sep 8, 2025
SHA
d47e7d13

Try a prompt

Base models

Task
Image-Text-to-Text
Library
transformers
Type
mllama
License
llama3.2
Languages
en
Created
May 26, 2025
Updated
Sep 8, 2025
Llama-3.2-11B-Vision-Instruct-StarFlow — AI Model — AIMarketly