RGBD20K: A Large-Scale Benchmark for RGB-D Semantic Segmentation Paper • 2609.29028 • Published 6 days ago • 10
DeltaWAM: Delta World Action Models for Bimanual Manipulation Paper • 2609.28811 • Published 7 days ago • 14
World Action Agent: Harnessing VLMs for Robot Manipulation via World Action Rehearsal Paper • 2609.29964 • Published 6 days ago • 12
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 7 days ago • 35
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs Paper • 2609.29845 • Published 6 days ago • 90
Coding Agents for Generalized Task and Motion Planning Problems Paper • 2609.30233 • Published 6 days ago • 19
HARMONY: Hierarchical Agentic Reasoning for MONocular Image-to-Scene Synthesis Paper • 2609.26793 • Published 8 days ago • 8
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published 23 days ago • 17
RoboFollow: Unveiling the Instruction Following Mirage in Embodied Agents Paper • 2609.25636 • Published 8 days ago • 13
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 9 days ago • 101
GraphSkillEvo: Evolutionary Optimization of Graph-Structured Agent Skills Paper • 2609.21749 • Published 12 days ago • 19
The Other Half of the Memory Wall: Serving 35B MoEs from SSD with Trained Routing Prediction Paper • 2609.18063 • Published 14 days ago • 19
CADWorld: Computer-Use Benchmark for Long-Horizon Computer-Aided Design Paper • 2609.16251 • Published 16 days ago • 14
All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts Paper • 2609.24058 • Published 9 days ago • 55
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 13 days ago • 75
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 9 days ago • 213
FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations Paper • 2609.20817 • Published 13 days ago • 40
TeleAntiFraud 2.0: A Refreshable, Profile-Grounded, and Audio-Based Benchmark for Telecom Fraud Detection Paper • 2609.18748 • Published 13 days ago • 10
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training Paper • 2609.26774 • Published 8 days ago • 56