Skip to content

sanps

fVLM-1.7B

sanps/fVLM-1.7B

fVLM-1.7B is a image-text-to-text model from sanps. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for pytorch. The card lists the license as apache-2.0.

A vision-language model that uses foveated attention to compress each video frame into a single visual token, enabling efficient processing of long videos on a single GPU.

Downloads · 30 days

7

8% of all-time downloads

All-time downloads

87

Public

Parameters

1.8B

14.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pt10.5 GB · 74%

Parameter types

How the weights are stored.

BF161.8B · 99%

Try a prompt

Task
Image-Text-to-Text
Library
pytorch
Type
foveated_vlm
License
apache-2.0
Languages
en
Created
Feb 22, 2026
Updated
Feb 24, 2026
fVLM-1.7B — AI Model — AIMarketly