RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning Paper • 2609.20784 • Published 3 days ago • 26
Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents Paper • 2609.17653 • Published 5 days ago • 25
PaperGym: Rubric-Centered Evolution for Research-Plan Generation Paper • 2608.31119 • Published 20 days ago • 33
Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning Paper • 2608.23318 • Published 27 days ago • 32
Embodied-Navigator: Point, Think, Memorize, and Align for Efficient Navigation Paper • 2608.17512 • Published Aug 18 • 52
Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text Paper • 2607.21072 • Published Jul 23 • 35
Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text Paper • 2607.21072 • Published Jul 23 • 35
VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Paper • 2607.01804 • Published Jul 2 • 27
GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification Paper • 2604.14258 • Published Apr 15 • 23
Visual Generation Unlocks Human-Like Reasoning through Multimodal World Models Paper • 2601.19834 • Published Jan 27 • 25
N3D-VLM: Native 3D Grounding Enables Accurate Spatial Reasoning in Vision-Language Models Paper • 2512.16561 • Published Dec 18, 2025 • 20