LLM-as-Jev: LLMs Are Already Jev-Style Decision Models -- When and How to Fine-Tune Them Paper • 2610.02076 • Published 4 days ago • 11
Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL Paper • 2609.20715 • Published 21 days ago • 44
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention Paper • 2609.15810 • Published 24 days ago • 51
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published Sep 7 • 376
CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation Paper • 2609.04083 • Published Sep 3 • 27
MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use Paper • 2608.20202 • Published Aug 20 • 34
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 266
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure Paper • 2608.11079 • Published Aug 11 • 17
TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex Paper • 2607.22143 • Published Aug 4 • 7
Knowledge-Geometry Decoupling: Refreshable Pretrained Transfer for Streaming Recommendation Paper • 2608.02738 • Published Aug 6 • 47