What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 9 days ago • 136
Chinese-Jev: Bringing System One Model to Chinese-Language Tasks Paper • 2609.36965 • Published 9 days ago • 25
Post-Training Leaves Behavioral Shadows on Unrelated Decisions Paper • 2609.29233 • Published 14 days ago • 273
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures Paper • 2609.29429 • Published 14 days ago • 29
Neural Spectral Capacity: Measuring and Designing Architectures from Network Specification Alone Paper • 2609.23087 • Published 19 days ago • 11
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 16 days ago • 164
Transferring the Intelligence of VLMs to Robotic Control Paper • 2609.22966 • Published 19 days ago • 120
EvoOntology: A Self-Evolving Ontology Layer for Data Agents Paper • 2609.15779 • Published 24 days ago • 159