MathSticks: A Benchmark for Visual Symbolic Compositional Reasoning with Matchstick Puzzles Paper • 2510.00483 • Published Oct 1, 2025
OneVision-Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence Paper • 2602.08683 • Published Feb 9 • 52
PRM-as-a-Judge: A Dense Evaluation Paradigm for Fine-Grained Robotic Auditing Paper • 2603.21669 • Published Mar 23 • 1
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence Paper • 2605.25979 • Published May 25 • 27
Towards Cross-View Point Correspondence in Vision-Language Models Paper • 2512.04686 • Published Dec 4, 2025
RoboOS-NeXT: A Unified Memory-based Framework for Lifelong, Scalable, and Robust Multi-Robot Collaboration Paper • 2510.26536 • Published Oct 30, 2025
Robo-Dopamine: General Process Reward Modeling for High-Precision Robotic Manipulation Paper • 2512.23703 • Published Dec 29, 2025 • 7
Robo-Dopamine: General Process Reward Modeling for High-Precision Robotic Manipulation Paper • 2512.23703 • Published Dec 29, 2025 • 7
LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training Paper • 2509.23661 • Published Sep 28, 2025 • 51
RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete Paper • 2502.21257 • Published Feb 28, 2025 • 2
Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning Paper • 2503.20752 • Published Mar 26, 2025 • 1
RoboOS: A Hierarchical Embodied Framework for Cross-Embodiment and Multi-Agent Collaboration Paper • 2505.03673 • Published May 6, 2025 • 2