On-Policy Parameter Update Direction Underlies Generalization in LLM Post-Training Paper • 2609.36659 • Published 8 days ago • 81
Video Generation Models: A Survey of Post-Training and Alignment Paper • 2610.00812 • Published 7 days ago • 60
A Missing Piece for Trustworthy AI Reviewers: From Benchmarking Rhetorical Robustness to SciCore Review Paper • 2609.39027 • Published 7 days ago • 76
Omni-IO Skills: Harnessing Your Agent Omni-Native Paper • 2609.31847 • Published 12 days ago • 448
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 10 days ago • 567
CoWindow Attention: Full Causal Coverage Is a Collective Property Paper • 2609.32704 • Published 11 days ago • 67
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published Sep 1 • 66
Running Featured 853 Agent Memory Leaderboard 🧠853 Unified memory evaluation · Results expected August 12.
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements Paper • 2608.17310 • Published Aug 18 • 110
How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review Paper • 2608.08975 • Published Aug 10 • 48
Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation Paper • 2607.05382 • Published Jul 9 • 83