Downloads · 30 days
523
49% of all-time downloads
prithivMLmods/Video-ORA-9B-GGUF
Video-ORA-9B-GGUF is a video-text-to-text model from prithivMLmods. Use it for the video-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
Video-ORA-9B is a 9-billion-parameter unified video understanding model built on Qwen3.5-9B and post-trained with OraRL (Annotations as Rollouts) — an annotation-augmented, on-policy reinforcement learning method that…
Downloads · 30 days
523
49% of all-time downloads
All-time downloads
1.1K
Public
Repo size
85.9 GB
Likes
2
Public
Click a slice to open those files.
.gguf86.9 GB · 100%
From the Hugging Face model README
Video-ORA-9B is a 9-billion-parameter unified video understanding model built on Qwen3.5-9B and post-trained with OraRL (Annotations as Rollouts) — an annotation-augmented, on-policy reinforcement learning method that equips a single model to handle seven task families with direct, task-native answers and no chain-of-thought decoding: temporal grounding, visual tracking, image/video segmentation, spatial grounding, spatial-temporal grounding, video question answering, and spatial intelligence. It retains a 262,144-token native context window inherited from its base model and leads matched seven-family benchmark comparisons against multimodal baselines despite skipping reasoning traces at inference time, with evaluations run using direct-answer prompts (
enable_thinking=False) to match its reported protocol. The model is served via vLLM (with aqwen3reasoning parser and configurable frame sampling for video input) or a Transformers serving endpoint, occupies roughly 17.6 GiB in BF16 weight loading, and is intended for research on structured video/spatial perception, benchmark evaluation, and task-specific adaptation — explicitly out of scope for safety-critical decisions, identity inference, or surveillance deployment. Trained on public dataset splits with evaluation identities and media excluded from the training mixture, it is released under the Apache License 2.0, consistent with its Qwen3.5-9B base.
[!NOTE] Model: https://huggingface.co/OraRL/Video-ORA-9B
| File Name | Quant Type | File Size | File Link |
|---|---|---|---|
| Video-ORA-9B.BF16.gguf | BF16 | 17.9 GB | Download |
| Video-ORA-9B.F16.gguf | F16 | 17.9 GB | Download |
| Video-ORA-9B.Q3_K_L.gguf | Q3_K_L | 4.93 GB | Download |
| Video-ORA-9B.Q3_K_M.gguf | Q3_K_M | 4.62 GB | Download |
| Video-ORA-9B.Q3_K_S.gguf | Q3_K_S | 4.26 GB | Download |
| Video-ORA-9B.Q4_0.gguf | Q4_0 | 5.31 GB | Download |
| Video-ORA-9B.Q4_K_M.gguf | Q4_K_M | 5.63 GB | Download |
| Video-ORA-9B.Q4_K_S.gguf | Q4_K_S | 5.35 GB | Download |
| Video-ORA-9B.Q5_0.gguf | Q5_0 | 6.31 GB | Download |
| Video-ORA-9B.Q5_K_M.gguf | Q5_K_M | 6.47 GB | Download |
| Video-ORA-9B.Q5_K_S.gguf | Q5_K_S | 6.31 GB | Download |
| Video-ORA-9B.mmproj-bf16.gguf | mmproj-bf16 | 922 MB | Download |
| Video-ORA-9B.mmproj-f16.gguf | mmproj-f16 | 922 MB | Download |
LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp