From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation Paper โข 2603.15600 โข Published 3 days ago โข 5 โข 3