Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published 8 days ago • 107
AGVBench: A Reliability-Oriented Benchmark of Data Augmentation for Vein Recognition Paper • 2607.02271 • Published 20 days ago • 17
Agent Explorative Policy Optimization for Multimodal Agentic Reasoning Paper • 2605.28774 • Published May 27 • 93
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433
DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards Paper • 2605.21467 • Published May 20 • 207
jeqcho/qwen-2.5-14b-instruct-all-animals-bottom_proj-lion-seed1 Text Generation • Updated May 23 • 3 • 1
ETHrobotlearning/smolvla_task2-color_from-relative2000_small2k-step1400 Robotics • 0.5B • Updated May 21 • 3 • 1