Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable Rewards Paper • 2609.03181 • Published 18 days ago • 8
The Router Within: Eliciting Native Skill Routing from a Frozen LLM Paper • 2609.15982 • Published 6 days ago • 6
view article Article One sandbox per rollout, or how labs run RL for agents in 2026 sergiopaniego • 8 days ago • 10
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published 10 days ago • 40
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models Paper • 2609.08418 • Published 12 days ago • 132
Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training Paper • 2609.15051 • Published 6 days ago • 12
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 16 days ago • 113
NeoMME: A Single-Tower Multimodal-Native Multilingual Foundation Encoder for Efficient Fine-Tuning and Inference Paper • 2609.01657 • Published 20 days ago • 34
NeoMME Collection Meet NeoMME: a family of 260M and 800M Multimodal-Native Multilingual Encoders • 12 items • Updated 16 days ago • 33
view article Article NeoMME: an efficient Multimodal-native and Multilingual Encoder Hcompany • 16 days ago • 108
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents Paper • 2609.09153 • Published 12 days ago • 40
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 17 days ago • 125
view article Article Training a coding model to paint watercolours with TRL and OpenEnv sergiopaniego • 17 days ago • 69
Kraken PP-OCRv6 text recognition models Collection Hub mirrors of Benjamin Kiessling's multilingual PP-OCRv6 line-recognition family for Kraken: tiny, small, and medium. • 3 items • Updated 16 days ago • 8