arxiv:2609.33848
Chujie Zheng
chujiezheng
AI & ML interests
Large Language Models
Recent Activity
authored a paper about 8 hours ago
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents authored a paper about 8 hours ago
Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models upvoted a paper about 9 hours ago
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents