Looping Beyond Twice: A Scalable Recipe for Looped Mixture-of-Experts Paper • 2610.01153 • Published 5 days ago • 15
Science Utopia? Closed-Loop LLM Simulation of Academic Research Ecosystems Paper • 2610.01257 • Published 5 days ago • 39
LexReward: A Taxonomy-Driven Reward Framework for Legal Language Models Paper • 2609.39071 • Published 6 days ago • 51
RealCompanion: Benchmarking Human Understanding from Reasoning over Longitudinal Real-World Conversations Paper • 2610.01780 • Published 5 days ago • 258
Pivot-SD: Efficient Self-Distillation for Masked Diffusion Language Models Paper • 2610.03665 • Published 4 days ago • 51
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 9 days ago • 564
Scaling Properties of Same-Family On-Policy Distillation Paper • 2609.32722 • Published 10 days ago • 323