Jinyang Wu
Jinyang23
AI & ML interests
large language models, reasoning, agentic rl
Recent Activity
authored a paper about 12 hours ago
Two-Stage Regularization-Based Structured Pruning for LLMs authored a paper about 12 hours ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning upvoted a paper about 17 hours ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement LearningOrganizations
None yet