Submitted by taesiri 16 Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills QwenBusinessUnit-Edu 0
Submitted by lhpku20010120 14 DataPrep-Bench: Benchmarking LLMs as Training Data Preparators · 14 authors 218 1
Submitted by taesiri 10 Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning NVIDIA 605
Submitted by hongst 6 LAMAR: An Open Language-Aware Multilingual Alignment Reranker NLP & AI - Korea University
Submitted by gdadhich 5 Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems · 1 authors 0 1
Submitted by AmirhoseinGH 4 Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making huawei technologies co. ltd 2 1
Submitted by soujanyaporia 3 IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation Deep Cognition and Language Research (DeCLaRe) Lab 1 1
Submitted by taesiri 2 Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering · 4 authors
Submitted by zpatrick 1 VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression · 4 authors 11 2