SAF-OPD: Stable Advantage Fusion for On-Policy Distillation Paper • 2607.29209 • Published 4 days ago • 27
MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling Paper • 2602.03359 • Published Feb 3 • 10
Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding Paper • 2509.06923 • Published Sep 8, 2025 • 22