yicheng qiu
MaXWe1l1
ยท
AI & ML interests
AI for science
RL theory
Recent Activity
liked a dataset about 1 month ago
SciDataOcean/ReasonEM upvoted a paper 3 months ago
Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO