arxiv:2608.01837
Jinyang Wu
Jinyang23
AI & ML interests
large language models, reasoning, agentic rl
Recent Activity
authored a paper about 19 hours ago
Two-Stage Regularization-Based Structured Pruning for LLMs authored a paper about 19 hours ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning upvoted a paper about 24 hours ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement LearningOrganizations
None yet