CalVerT: Augmenting Agents with Calibrated Verifier Telemetry Improves Action and Learning in Knowledge-Intensive Tasks Paper • 2606.21777 • Published Jun 19 • 5
No Hidden Prompts Needed! You Can Game AI Peer Review with Presentation-Only Revisions Paper • 2606.13044 • Published Jun 11 • 11
Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems Paper • 2605.26302 • Published May 25 • 29
TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning Paper • 2603.12529 • Published Mar 13 • 19
DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset Paper • 2403.12945 • Published Mar 19, 2024 • 2
Mini-BEHAVIOR: A Procedurally Generated Benchmark for Long-horizon Decision-Making in Embodied AI Paper • 2310.01824 • Published Oct 3, 2023 • 1
Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning Paper • 2603.11653 • Published Mar 12 • 2
EntRGi: Entropy Aware Reward Guidance for Diffusion Language Models Paper • 2602.05000 • Published Feb 4 • 2
Mitigating Catastrophic Forgetting in Mathematical Reasoning Finetuning through Mixed Training Paper • 2512.13706 • Published Dec 5, 2025 • 1
Sasha: Creative Goal-Oriented Reasoning in Smart Homes with Large Language Models Paper • 2305.09802 • Published May 16, 2023 • 1
Flavors of Moonshine: Tiny Specialized ASR Models for Edge Devices Paper • 2509.02523 • Published Sep 2, 2025 • 22