SED: Self-Evaluation Decoding Enhances Large Language Models for Better Generation Paper • 2405.16552 • Published May 26, 2024 • 1
Mind the Generation Process: Fine-Grained Confidence Estimation During LLM Generation Paper • 2508.12040 • Published Aug 16, 2025 • 14
Selective Expert Guidance for Effective and Diverse Exploration in Reinforcement Learning of LLMs Paper • 2510.04140 • Published Oct 5, 2025
GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0) Paper • 2604.17091 • Published Apr 18 • 25
ADaPT: Token-Level Decoupling for Efficient Large Reasoning Models Paper • 2606.19919 • Published Jun 18
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 7 days ago • 1
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 7 days ago • 1
Mind the Generation Process: Fine-Grained Confidence Estimation During LLM Generation Paper • 2508.12040 • Published Aug 16, 2025 • 14