LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published 21 days ago • 14
Fathom: Per-Query Read Depth for Sparse Decoding over Offloaded KV Caches Paper • 2609.17652 • Published 13 days ago • 8
open-llm-leaderboard-old/details_xformAI__opt-125m-gqa-ub-6-best-for-KV-cache Updated Jan 23, 2024 • 394 • 3
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 11 days ago • 184
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 7 days ago • 96
Complex KDA: Understanding and Enhancing the Expressivity of Kimi Delta Attention Paper • 2609.24797 • Published 7 days ago • 11
onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction Paper • 2609.24983 • Published 7 days ago • 55