Beyond Token-level Supervision: Unlocking the Potential of Decoding-based Regression via Reinforcement Learning Paper • 2512.06533 • Published Dec 6, 2025 • 9
PVChat: Personalized Video Chat with One-Shot Learning Paper • 2503.17069 • Published Mar 21, 2025 • 8
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System Paper • 2511.17943 • Published Nov 22, 2025 • 4
LaS-Comp: Zero-shot 3D Completion with Latent-Spatial Consistency Paper • 2602.18735 • Published Feb 21 • 2
One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems Paper • 2605.22144 • Published May 21 • 11
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL Paper • 2605.18703 • Published May 18 • 51
CHARM: Calibrating Reward Models With Chatbot Arena Scores Paper • 2504.10045 • Published Apr 14, 2025 • 1
CodeScaler: Scaling Code LLM Training and Test-Time Inference via Execution-Free Reward Models Paper • 2602.17684 • Published Feb 4 • 23
Beyond Message Passing: A Semantic View of Agent Communication Protocols Paper • 2604.02369 • Published Apr 13 • 1
Training Diffusion Language Models for Black-Box Optimization Paper • 2603.17919 • Published May 29 • 12
Discrete Diffusion Models: A Unified Framework from Tokenization to Generation Paper • 2607.13431 • Published 24 days ago • 20
Beyond Message Passing: A Semantic View of Agent Communication Protocols Paper • 2604.02369 • Published Apr 13 • 1