ActionPiece: Rethinking Action Tokenization for Autoregressive Vision-Language-Action Models Paper • 2609.18487 • Published 5 days ago • 41
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 4 days ago • 94
An Empirical Study of Harness Design for Coding Agents Paper • 2609.20804 • Published 4 days ago • 72
BoldingBuilds/Ternary-Bonsai-2-27B-Abliterated-PTQ1_0-GGUF Text Generation • 27B • Updated 3 days ago • 16.9k • 42
Disentangling Representation Evolution in Transformers through Directional Decomposition Paper • 2609.15975 • Published 7 days ago • 9
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 5 days ago • 65
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models Paper • 2609.08418 • Published 13 days ago • 135
DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published 16 days ago • 163
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 17 days ago • 114
Discovery Foundation Models: Toward Open-Ended Discovery Intelligence Paper • 2609.15973 • Published 7 days ago • 33
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF Image-Text-to-Text • 177B • Updated 3 days ago • 43k • 195
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 7 days ago • 239
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published 11 days ago • 44