Skip to content

vllm-sr

multi-modal-embed-small

vllm-sr/multi-modal-embed-small

multi-modal-embed-small is a sentence similarity model from vllm-sr. Use it when you need a score for how close two texts are. It is set up for transformers. The card lists the license as apache-2.0.

A compact multimodal embedding model that unifies text, image, and audio representations in a shared semantic space. Part of the MoM (Mixture of Models) family.

Downloads · 30 days

6.5K

33% of all-time downloads

All-time downloads

19.6K

Public

Parameters

338M

5.2 GB on disk

Likes

27

Trending 1

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors1.4 GB · 48%

At a glance

Task
Sentence Similarity
Library
transformers
License
apache-2.0
Access
Public
Created
Feb 5, 2026
Updated
Oct 2, 2026
SHA
cba24df2
Task
Sentence Similarity
Library
transformers
License
apache-2.0
Languages
en
Created
Feb 5, 2026
Updated
Oct 2, 2026
multi-modal-embed-small — AI Model — AIMarketly