Cross-Lingual Alignment for Decoder-Only Models using MoE Routers Paper • 2610.01921 • Published 6 days ago • 2
VDiff-Bench: A Challenging Benchmark for Fine-Grained Image Difference Identification Paper • 2609.06245 • Published Sep 5 • 28
Stream4D: 4D-Consistency for Streaming Autoregressive Diffusion Video Models Paper • 2608.19556 • Published Aug 20 • 5
Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care Paper • 2606.31036 • Published Jun 30 • 6
HarnessBridge: Learnable Bidirectional Controller for LLM Agent Harness Paper • 2606.12882 • Published Jun 11 • 15
Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization Paper • 2605.29198 • Published May 29 • 2
Less is More: Early Stopping Rollout for On-Policy Distillation Paper • 2605.27028 • Published May 26 • 12
ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning Paper • 2602.21534 • Published Feb 25 • 26
Adam Improves Muon: Adaptive Moment Estimation with Orthogonalized Momentum Paper • 2602.17080 • Published Feb 19 • 3
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments Paper • 2602.02459 • Published Feb 2 • 4