ActionPiece: Rethinking Action Tokenization for Autoregressive Vision-Language-Action Models Paper • 2609.18487 • Published 6 days ago • 44
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 5 days ago • 97
An Empirical Study of Harness Design for Coding Agents Paper • 2609.20804 • Published 5 days ago • 74
BoldingBuilds/Ternary-Bonsai-2-27B-Abliterated-PTQ1_0-GGUF Text Generation • 27B • Updated 3 days ago • 20.8k • 47
Disentangling Representation Evolution in Transformers through Directional Decomposition Paper • 2609.15975 • Published 8 days ago • 9
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 6 days ago • 73
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models Paper • 2609.08418 • Published 14 days ago • 135
DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published 17 days ago • 163
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 18 days ago • 116
Discovery Foundation Models: Toward Open-Ended Discovery Intelligence Paper • 2609.15973 • Published 8 days ago • 33
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF Image-Text-to-Text • 177B • Updated 3 days ago • 53.1k • 210
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 8 days ago • 241
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published 12 days ago • 45