iFAN: Inference-Aware Learning for Plain Mask Transformers Paper • 2608.03216 • Published 8 days ago • 9
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 23 days ago • 154
From Pixels to States: Rethinking Interactive World Models as Game Engines Paper • 2607.14076 • Published about 1 month ago • 37
TerraDiT-Ω: Unified Spatial Control for Satellite Image Synthesis with Any Geospatial Primitive Paper • 2606.31029 • Published Jun 30 • 6
OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data Paper • 2606.13432 • Published Jun 11 • 113
YoCausal: How Far is Video Generation from World Model? A Causality Perspective Paper • 2605.30346 • Published May 28 • 56
MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation Paper • 2606.02470 • Published Jun 1 • 16
IndusAgent: Reinforcing Open-Vocabulary Industrial Anomaly Detection with Agentic Tools Paper • 2605.20682 • Published May 20 • 86
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining Paper • 2605.14747 • Published May 14 • 147
SlimSpec: Low-Rank Draft LM-Head for Accelerated Speculative Decoding Paper • 2605.10453 • Published May 11 • 9
LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model Paper • 2604.20796 • Published Apr 22 • 243
DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models Paper • 2603.26164 • Published Mar 27 • 366
Consistency Amplifies: How Behavioral Variance Shapes Agent Accuracy Paper • 2603.25764 • Published Mar 26 • 5
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning Paper • 2604.02721 • Published Apr 3 • 639
ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers Paper • 2603.24414 • Published Mar 25 • 183