Jinyang Wu
Jinyang23
AI & ML interests
large language models, reasoning, agentic rl
Recent Activity
authored a paper about 11 hours ago
Two-Stage Regularization-Based Structured Pruning for LLMs authored a paper about 11 hours ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning upvoted a paper about 16 hours ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement LearningOrganizations
None yet