Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning
xhl
zjuxhl
AI & ML interests
None yet
Recent Activity
upvoted a paper about 20 hours ago
Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents upvoted a paper about 20 hours ago
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning updated a Space 9 days ago
zjuxhl/EasySteer