Đặng Tuấn
hcmnguyen
·
AI & ML interests
Coffee addict. Dog person.
Recent Activity
upvoted a paper about 12 hours ago
TRACE: Rollout-Guided Quantization-Aware Training for FP4 Reinforcement Learning of MoE Language Models upvoted a paper about 12 hours ago
LoGRA: Scaling LLM Reinforcement Learning with Low-Rank Gradient Sketches upvoted a paper about 12 hours ago
False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search AgentsOrganizations
None yet