Skip to content

flow666

MLLM-4D

flow666/MLLM-4D

MLLM-4D is a video-text-to-text model from flow666. Use it for the video-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers.

MLLM-4D is a comprehensive framework designed to bridge the gaps in training data curation and model post-training for spatiotemporal understanding and reasoning. It enables multimodal large language models (MLLMs) to…

Downloads · 30 days

0

Access

Public

Updated Mar 3, 2026

Repo size

35.1 GB

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors35.1 GB · 100%

At a glance

Task
Video-Text-to-Text
Library
transformers
Access
Public
Created
Feb 3, 2026
Updated
Mar 3, 2026
SHA
16bb0c57
Task
Video-Text-to-Text
Library
transformers
Created
Feb 3, 2026
Updated
Mar 3, 2026