Skip to content

atlas-institute

code-trainer-vision-adapter

atlas-institute/code-trainer-vision-adapter

code-trainer-vision-adapter is a image-to-text model from atlas-institute. Use it when you need a caption or text from an image. It is set up for peft. The card lists the license as apache-2.0.

A multimodal screenshot → code model: a frozen Swin-B vision encoder, an MLP projector, and a LoRA adapter for Qwen/Qwen2.5-Coder-1.5B-Instruct.

Downloads · 30 days

0

Access

Public

Updated Aug 12, 2026

Repo size

80.5 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors73.9 MB · 92%

At a glance

Task
Image-to-Text
Library
peft
License
apache-2.0
Access
Public
Created
May 2, 2026
Updated
Aug 12, 2026
SHA
9bd31d42

Base models

Task
Image-to-Text
Library
peft
License
apache-2.0
Created
May 2, 2026
Updated
Aug 12, 2026