Retrieval-Augmented Skill Optimization via Cross-Harness Adaptation Paper • 2609.38024 • Published 4 days ago • 41
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published about 1 month ago • 104
Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation Paper • 2609.08798 • Published 25 days ago • 84
Reason in the Words You Speak: Idiolectal Paraphrasing Off-Policy Traces for Reasoning Distillation in VideoLLMs Paper • 2608.26684 • Published Aug 27 • 24
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data Paper • 2609.05405 • Published 29 days ago • 45
AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems Paper • 2609.08572 • Published 25 days ago • 108
Retrieve What's Missing: Coverage-Maximizing Retrieval for Consistent Long Video Generation Paper • 2606.02479 • Published Jun 1 • 25
F4Splat: Feed-Forward Predictive Densification for Feed-Forward 3D Gaussian Splatting Paper • 2603.21304 • Published Mar 22 • 34
Bidirectional Likelihood Estimation with Multi-Modal Large Language Models for Text-Video Retrieval Paper • 2507.23284 • Published Jul 31, 2025 • 4
Representation Shift: Unifying Token Compression with FlashAttention Paper • 2508.00367 • Published Aug 1, 2025 • 16
LLaMo: Large Language Model-based Molecular Graph Assistant Paper • 2411.00871 • Published Oct 31, 2024 • 22