Downloads · 30 days
161
13% of all-time downloads
zhang9302002/Flash-VStream-Qwen-7b
Flash-VStream-Qwen-7b is a machine learning model from zhang9302002. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
<a href='https://zhang9302002.github.io/vstream-iccv-page/'<img src='https://img.shields.io/badge/Project-Page-Green'</a <a href='https://arxiv.org/abs/2506.23825'<img src='https://img.shields.io/badge/Paper-Arxiv-red…
Downloads · 30 days
161
13% of all-time downloads
All-time downloads
1.2K
Public
Parameters
8.3B
16.6 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors16.6 GB · 100%
From the Hugging Face model README
<a href='https://zhang9302002.github.io/vstream-iccv-page/'><img src='https://img.shields.io/badge/Project-Page-Green'></a> <a href='https://arxiv.org/abs/2506.23825'><img src='https://img.shields.io/badge/Paper-Arxiv-red'></a> <a href="https://github.com/IVGSZ/Flash-VStream"><img src='https://img.shields.io/badge/Github-Code-blue'></a>
We proposed Flash-VStream, an efficient VLM with a novel Flash Memory mechanism that enables real-time understanding and Q&A of extremely long video streams. Our model achieves outstanding accuracy and efficiency on EgoSchema, MLVU, LVBench, MVBench and Video-MME Benchmarks.
This project is licensed under the Apache 2.0 License.