DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models Paper • 2503.04472 • Published Jan 12
HiMPO: Hindsight-Informed Memory Policy Optimization for Less-Entangled Credit in Long-Horizon Agents Paper • 2606.16285 • Published Jun 15 • 1
HEAL: Hindsight Entropy-Assisted Learning for Reasoning Distillation Paper • 2603.10359 • Published Mar 11
AuraFusion360: Augmented Unseen Region Alignment for Reference-based 360° Unbounded Scene Inpainting Paper • 2502.05176 • Published Feb 7, 2025 • 40