Selection-Based Structured Reasoning: Toward Efficient Multimodal Search Agents Paper • 2610.01892 • Published 10 days ago • 30
Rethinking Cross-Tokenizer On-Policy Distillation: From Alignment Coverage to Supervision Reliability Paper • 2610.08448 • Published 5 days ago • 191
QuantWM: Temporally Consistent 2-Bit KV Cache Quantization for Video World Models Paper • 2609.26425 • Published 13 days ago • 19
Native Action-Prior Learning from Videos for World Action Models Paper • 2610.03391 • Published 9 days ago • 88
Knowing When Thinking Is Not Enough: Teaching Small Reasoning Models to Reason Beyond Their Parametric Knowledge Paper • 2609.34327 • Published 13 days ago • 41
InternW0-Δ: A World Action Model Bridging Predictive Dynamics and Actions with 20K+ Hours of Open Data Paper • 2609.31394 • Published 16 days ago • 31
FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation Paper • 2609.27657 • Published 18 days ago • 9
The Past Frames the Future: Memory for Autoregressive Video Generation Paper • 2609.28466 • Published 18 days ago • 65
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published Sep 7 • 17
Fathom: Per-Query Read Depth for Sparse Decoding over Offloaded KV Caches Paper • 2609.17652 • Published 26 days ago • 8
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 24 days ago • 228
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 20 days ago • 104
Complex KDA: Understanding and Enhancing the Expressivity of Kimi Delta Attention Paper • 2609.24797 • Published 20 days ago • 11
onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction Paper • 2609.24983 • Published 20 days ago • 55