Downloads · 30 days
14
16% of all-time downloads
tsinghua-ee/video_SALMONN2plus_3B_audioAlign
video_SALMONN2plus_3B_audioAlign is a machine learning model from tsinghua-ee. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
Audio aligned model of video-SALMONN 2+ 3B model
Downloads · 30 days
14
16% of all-time downloads
All-time downloads
90
Public
Parameters
5B
10 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors10 GB · 100%
How the weights are stored.
BF165B · 100%
From the Hugging Face model README
Audio aligned model of video-SALMONN 2+ 3B model