PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning Paper • 2608.01837 • Published 3 days ago • 37
Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models Paper • 2503.02318 • Published Mar 4, 2025 • 3
CompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon Agents Paper • 2607.05378 • Published Jul 6 • 1
Code2Logic: Game-Code-Driven Data Synthesis for Enhancing VLMs General Reasoning Paper • 2505.13886 • Published May 20, 2025 • 10
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 7 days ago • 53