StyleForge: Indoor Furniture Styling by Counterfactual Reasoning in a Hypergraph Field Paper • 2608.01954 • Published 2 days ago • 12
TARS: Timestep-Aware Data Scaling for 3D-Free Video Re-Shooting Paper • 2607.28261 • Published 6 days ago • 116
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 13 days ago • 151
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170