Video-HopChain-8B-Standard-RL-GGUF

Video-HopChain-8B-Standard-RL is the stage-1 checkpoint of the Video-HopChain project, a Qwen3-VL-8B-Instruct model trained with GRPO on a 105,993-row general video QA mixture (LLaVA-Video, STAR, CLEVRER, NExT-QA, and PerceptionTest) at 24 frames, saved at step 80 — the "+ standard RL" row of Table 1 in the paper "Video-HopChain: Multi-Hop Questions and Confidence-Gated Exploration for Video Reasoning Models." It saw no Video-HopChain data and used no Confidence-Gated Exploration, serving instead as the common starting point for both second-stage runs, including the final ngqtrung/Video-HopChain-8B, as well as a baseline for reproducing Table 1. Evaluated at 100 frames per video across eight benchmarks, it improves over the unmodified Qwen3-VL-8B-Instruct base — 55.4% mean accuracy versus 52.3%, with notable gains on PerceptionComp (34.3% vs 28.1%) and Video-Holmes (47.4% vs 40.7%) — though it shows no improvement on the in-domain held-out Video-HopChain split (13.4% for both), underscoring that general video RL alone does not transfer to the multi-hop reasoning task the full pipeline targets. The model reasons inside <think> tags and outputs a boxed final answer, is loadable via standard Transformers (Qwen3VLForConditionalGeneration) in BF16, and is released under the Apache 2.0 license.

Model Files

File Name Quant Type File Size File Link Description
Video-HopChain-8B-Standard-RL.BF16.gguf BF16 16.4 GB Link Full BF16 weights. Highest quality, largest file size.
Video-HopChain-8B-Standard-RL.Q3_K_L.gguf Q3_K_L 4.43 GB Link Lower quality but usable, good for low RAM availability.
Video-HopChain-8B-Standard-RL.Q3_K_M.gguf Q3_K_M 4.12 GB Link Low quality.
Video-HopChain-8B-Standard-RL.Q4_K_M.gguf Q4_K_M 5.03 GB Link Good quality, default size for most use cases, recommended.
Video-HopChain-8B-Standard-RL.Q4_K_S.gguf Q4_K_S 4.8 GB Link Slightly lower quality with more space savings, recommended.
Video-HopChain-8B-Standard-RL.Q5_K_M.gguf Q5_K_M 5.85 GB Link High quality, recommended.
Video-HopChain-8B-Standard-RL.Q5_K_S.gguf Q5_K_S 5.72 GB Link High quality, recommended.
Video-HopChain-8B-Standard-RL.mmproj-bf16.gguf mmproj-bf16 1.16 GB Link Multimodal projection file in BF16 format. Used for vision/language models.

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Downloads last month
529
GGUF
Model size
8B params
Architecture
qwen3vl
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for prithivMLmods/Video-HopChain-8B-Standard-RL-GGUF

Quantized
(1)
this model

Collection including prithivMLmods/Video-HopChain-8B-Standard-RL-GGUF