Video-ORA-9B-GGUF

Video-ORA-9B is a 9-billion-parameter unified video understanding model built on Qwen3.5-9B and post-trained with OraRL (Annotations as Rollouts) — an annotation-augmented, on-policy reinforcement learning method that equips a single model to handle seven task families with direct, task-native answers and no chain-of-thought decoding: temporal grounding, visual tracking, image/video segmentation, spatial grounding, spatial-temporal grounding, video question answering, and spatial intelligence. It retains a 262,144-token native context window inherited from its base model and leads matched seven-family benchmark comparisons against multimodal baselines despite skipping reasoning traces at inference time, with evaluations run using direct-answer prompts (enable_thinking=False) to match its reported protocol. The model is served via vLLM (with a qwen3 reasoning parser and configurable frame sampling for video input) or a Transformers serving endpoint, occupies roughly 17.6 GiB in BF16 weight loading, and is intended for research on structured video/spatial perception, benchmark evaluation, and task-specific adaptation — explicitly out of scope for safety-critical decisions, identity inference, or surveillance deployment. Trained on public dataset splits with evaluation identities and media excluded from the training mixture, it is released under the Apache License 2.0, consistent with its Qwen3.5-9B base.

Model: https://huggingface.co/OraRL/Video-ORA-9B

Model Files

File Name Quant Type File Size File Link
Video-ORA-9B.BF16.gguf BF16 17.9 GB Download
Video-ORA-9B.F16.gguf F16 17.9 GB Download
Video-ORA-9B.Q3_K_L.gguf Q3_K_L 4.93 GB Download
Video-ORA-9B.Q3_K_M.gguf Q3_K_M 4.62 GB Download
Video-ORA-9B.Q3_K_S.gguf Q3_K_S 4.26 GB Download
Video-ORA-9B.Q4_0.gguf Q4_0 5.31 GB Download
Video-ORA-9B.Q4_K_M.gguf Q4_K_M 5.63 GB Download
Video-ORA-9B.Q4_K_S.gguf Q4_K_S 5.35 GB Download
Video-ORA-9B.Q5_0.gguf Q5_0 6.31 GB Download
Video-ORA-9B.Q5_K_M.gguf Q5_K_M 6.47 GB Download
Video-ORA-9B.Q5_K_S.gguf Q5_K_S 6.31 GB Download
Video-ORA-9B.mmproj-bf16.gguf mmproj-bf16 922 MB Download
Video-ORA-9B.mmproj-f16.gguf mmproj-f16 922 MB Download

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Downloads last month
469
GGUF
Model size
9B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for prithivMLmods/Video-ORA-9B-GGUF

Finetuned
Qwen/Qwen3.5-9B
Quantized
(3)
this model

Collection including prithivMLmods/Video-ORA-9B-GGUF