PReCache: Efficient KV Cache Sharing for Multi-LoRA Agents via Low-Rank Precomputation and Neutral Reconstruction Paper • 2609.34054 • Published 13 days ago • 6
KVCMAS: Efficient KV cache Correction for Shared Context in Multi-Agent Systems Paper • 2609.34060 • Published 13 days ago • 6
WaveFront Decoding: Parallelized Self-Speculative Decoding for Looped Language Models Paper • 2609.23033 • Published 15 days ago • 9