Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning Paper • 2606.18974 • Published Jun 17 • 3
Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems Paper • 2605.14892 • Published May 14 • 51
$\textbf{AGT$^{AO}$}$: Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality Paper • 2602.01703 • Published Feb 2 • 2
AGT^{AO}: Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality Paper • 2602.01703 • Published Feb 2 • 2
Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning Paper • 2606.18974 • Published Jun 17 • 3
Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems Paper • 2605.14892 • Published May 14 • 51
AERO: Autonomous Evolutionary Reasoning Optimization via Endogenous Dual-Loop Feedback Paper • 2602.03084 • Published Feb 3 • 1
OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions Paper • 2602.05843 • Published Feb 5 • 61
TIDE: Trajectory-based Diagnostic Evaluation of Test-Time Improvement in LLM Agents Paper • 2602.02196 • Published Feb 2 • 35
A^3-Bench: Benchmarking Memory-Driven Scientific Reasoning via Anchor and Attractor Activation Paper • 2601.09274 • Published Jan 14 • 84
Deliberation on Priors: Trustworthy Reasoning of Large Language Models on Knowledge Graphs Paper • 2505.15210 • Published May 21, 2025 • 19
Deliberation on Priors: Trustworthy Reasoning of Large Language Models on Knowledge Graphs Paper • 2505.15210 • Published May 21, 2025 • 19