Video-HopChain: Multi-Hop Questions and Confidence-Gated Exploration for Video Reasoning Models Paper • 2609.25773 • Published 4 days ago • 1
Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs Paper • 2605.23975 • Published May 13
Senses Wide Shut: A Representation-Action Gap in Omnimodal LLMs Paper • 2605.13737 • Published May 13
Video-HopChain: Multi-Hop Questions and Confidence-Gated Exploration for Video Reasoning Models Paper • 2609.25773 • Published 4 days ago • 1
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook 📚 3.31k The secrets to building world-class LLMs
LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training Paper • 2509.23661 • Published Sep 28, 2025 • 52
Video RLVR — final training data Collection The datasets behind our Qwen3-VL-8B video RLVR runs: the base 24f/100k mixture and the HopChain v5 multi-hop corpora. • 2 items • Updated 23 days ago