Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation
Paper • 2607.27372 • Published • 17
None defined yet.
AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents
SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning