Skip to content

tsinghua-ee

video_SALMONN2plus_3B_audioAlign

tsinghua-ee/video_SALMONN2plus_3B_audioAlign

video_SALMONN2plus_3B_audioAlign is a machine learning model from tsinghua-ee. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.

Audio aligned model of video-SALMONN 2+ 3B model

Downloads · 30 days

14

16% of all-time downloads

All-time downloads

90

Public

Parameters

5B

10 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors10 GB · 100%

Parameter types

How the weights are stored.

BF165B · 100%

Base models

Type
qwen2_5_vl
License
apache-2.0
Created
Feb 15, 2026
Updated
Feb 15, 2026