arxiv:2606.03197
Ziheng Li
ChillingDream
AI & ML interests
Natural Language Processing
Recent Activity
upvoted a paper about 24 hours ago
SAF-OPD: Stable Advantage Fusion for On-Policy Distillation authored a paper 2 months ago
To Mix or To Merge: Toward Multi-Domain Reinforcement Learning for Large Language Models authored a paper 2 months ago
LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant TasksOrganizations
None yet