Collections
Discover the best community collections!
Collections trending this week
-
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
Paper • 2607.14777 • Published • 106 -
Self-Improvements in Modern Agentic Systems: A Survey
Paper • 2607.13104 • Published • 33 -
Loop the Loopies!
Paper • 2607.16051 • Published • 77 -
When Does Muon Help Agentic Reinforcement Learning?
Paper • 2607.16169 • Published • 15
-
ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration
Paper • 2605.03042 • Published • 147 -
OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis
Paper • 2603.20278 • Published • 102 -
S1-Omni: A Unified Multimodal Reasoning Model for Scientific Understanding, Prediction, and Generation
Paper • 2607.15686 • Published • 16 -
DSWorld: A Data Science World Model for Efficient Autonomous Agents
Paper • 2607.15901 • Published • 12
-
ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration
Paper • 2605.03042 • Published • 147 -
OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis
Paper • 2603.20278 • Published • 102 -
S1-Omni: A Unified Multimodal Reasoning Model for Scientific Understanding, Prediction, and Generation
Paper • 2607.15686 • Published • 16 -
DSWorld: A Data Science World Model for Efficient Autonomous Agents
Paper • 2607.15901 • Published • 12
-
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
Paper • 2607.14777 • Published • 106 -
Self-Improvements in Modern Agentic Systems: A Survey
Paper • 2607.13104 • Published • 33 -
Loop the Loopies!
Paper • 2607.16051 • Published • 77 -
When Does Muon Help Agentic Reinforcement Learning?
Paper • 2607.16169 • Published • 15