MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities Paper • 2607.25948 • Published 3 days ago • 13
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 8 days ago • 149
Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning Paper • 2607.07708 • Published 23 days ago • 88
APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies Paper • 2606.12366 • Published Jun 10 • 5
LVSA: Training-Free Sparse Attention for Long Video Diffusion Paper • 2605.31057 • Published May 29 • 14
timm/mobilenetv3_small_100.lamb_in1k Image Classification • 2.55M • Updated Oct 19, 2025 • 18.6M • 99
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale Paper • 2605.27235 • Published May 26 • 9
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration Paper • 2605.20025 • Published May 19 • 191
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining Paper • 2605.14747 • Published May 14 • 147