MemoryCPT: An End-to-End Agent Memory Framework for Cost-Performance Trade-off Paper • 2608.04843 • Published Aug 5 • 3
The Personalization Mirage: How LLMs Fabricate User Profiles, and Why Self-Monitoring Misleads Paper • 2608.04570 • Published Aug 5 • 41
When Memory Lies: An Empirical Study of Spatial Memory Staleness in VLM Agents Paper • 2608.04574 • Published Aug 5 • 16
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adaptation in Coding Assistants Paper • 2607.26611 • Published Jul 29 • 33
Are LLMs Ready for Scientific Discovery? A Capability-Oriented Benchmark for AI Scientists Paper • 2607.11079 • Published Jul 13 • 5
Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering Paper • 2603.28583 • Published Jul 14 • 10
When Classic Cache Policies Fail: Learning-Augmented Replacement for Semantic Retrieval Buffers Paper • 2607.00394 • Published Jul 1 • 5
GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory Paper • 2605.01688 • Published May 3 • 2
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid? Paper • 2605.06527 • Published May 7 • 48
Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context Paper • 2605.13831 • Published May 13 • 90