SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent Paper • 2608.07449 • Published 8 days ago • 1
Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning Paper • 2601.13942 • Published Jan 20 • 2
Pushing the Boundaries of Natural Reasoning: Interleaved Bonus from Formal-Logic Verification Paper • 2601.22642 • Published Jan 30 • 9