Questioning the Questions: Sustaining Self-Evolution in Reasoning Models Paper • 2610.04299 • Published 5 days ago • 46
What Gradients Add to Text Leakage in Split Language Models, Counted per Token and per Document Paper • 2610.04128 • Published 6 days ago • 9
Pretrain Once, Route Anywhere: Towards a Foundation Model for LLM Routing Paper • 2609.37362 • Published 9 days ago • 17
What Does Privileged Information Add to On-Policy Self-Distillation? Paper • 2609.20612 • Published 21 days ago • 36
SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution Paper • 2609.05594 • Published Sep 4 • 36
Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM Paper • 2609.04098 • Published Sep 3 • 86
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published Sep 2 • 408
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published Aug 18 • 161
Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL Paper • 2608.17253 • Published Aug 19 • 100
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published Aug 14 • 287