arxiv:2605.15726
Chanuk Lee
tally0818
AI & ML interests
LLM post-training
Recent Activity
upvoted a paper about 13 hours ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement upvoted a paper about 13 hours ago
Motif 3: Technical Report liked a dataset about 13 hours ago
ArtificialAnalysis/AA-Omniscience-PublicOrganizations
None yet