pangpangxuan
pangxuan
AI & ML interests
None yet
Recent Activity
upvoted a paper about 16 hours ago
1% of Tokens Can Be Enough: On Gradient Estimation in On-Policy Distillation upvoted a paper 1 day ago
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay upvoted a paper 2 days ago
RRSI: Regularized Recursive Self-Improvement of Agent HarnessesOrganizations
None yet