Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts Paper • 2602.03473 • Published May 8 • 11
From Data to Rewards: a Bilevel Optimization Perspective on Maximum Likelihood Estimation Paper • 2510.07624 • Published Oct 8, 2025 • 9
FocuSFT: Bilevel Optimization for Dilution-Aware Long-Context Fine-Tuning Paper • 2605.09932 • Published May 11 • 2
Instance-Aware Parameter Configuration in Bilevel Late Acceptance Hill Climbing for the Electric Capacitated Vehicle Routing Problem Paper • 2605.00572 • Published May 1 • 1
A Tale of Two Problems: Multi-Task Bilevel Learning Meets Equality Constrained Multi-Objective Optimization Paper • 2605.09094 • Published May 14 • 1
DPOT: Auto-Regressive Denoising Operator Transformer for Large-Scale PDE Pre-Training Paper • 2403.03542 • Published Mar 6, 2024 • 1
Scorio.jl: A Julia package for ranking stochastic responses Paper • 2603.14103 • Published Mar 14 • 1
Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver Paper • 2604.25067 • Published Apr 29 • 2
OpenDevin: An Open Platform for AI Software Developers as Generalist Agents Paper • 2407.16741 • Published Jul 23, 2024 • 89
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 2 days ago • 113
AI-Trader: Benchmarking Autonomous Agents in Real-Time Financial Markets Paper • 2512.10971 • Published Dec 1, 2025 • 13
SkillOpt: Executive Strategy for Self-Evolving Agent Skills Paper • 2605.23904 • Published May 22 • 264
Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents Paper • 2609.23986 • Published 2 days ago • 14
ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration Paper • 2605.03042 • Published May 4 • 152