Revisiting Complete Reasoning Traces for Post-Training Paper ⢠2609.07103 ⢠Published 23 days ago ⢠23
Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation Paper ⢠2609.08798 ⢠Published 22 days ago ⢠84
Beneath the Surface of Chains-of-Thought: A Mechanistic Interpretation of Reasoning Operations in LLMs Paper ⢠2609.04753 ⢠Published 26 days ago ⢠16
Verification-Aware Training for Speculative Decoding Paper ⢠2608.30135 ⢠Published 30 days ago ⢠11
On-Policy Delta Distillation for Multilingual Math Reasoning Paper ⢠2608.05802 ⢠Published Aug 6 ⢠33
MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model Paper ⢠2602.06393 ⢠Published Feb 6 ⢠5
Weak-to-Strong Generalization via Direct On-Policy Distillation Paper ⢠2607.05394 ⢠Published Jul 8 ⢠144
Retrieve, Don't Retrain: Extending Vision Language Action Models to New Tasks at Test Time Paper ⢠2606.15631 ⢠Published Jun 14 ⢠17
MuCo Collection MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model [CVPR 2026] ⢠4 items ⢠Updated Apr 13 ⢠2
Grounding World Simulation Models in a Real-World Metropolis Paper ⢠2603.15583 ⢠Published Mar 16 ⢠154
Exploring Conditions for Diffusion models in Robotic Control Paper ⢠2510.15510 ⢠Published Oct 17, 2025 ⢠40
Map the Flow: Revealing Hidden Pathways of Information in VideoLLMs Paper ⢠2510.13251 ⢠Published Oct 15, 2025 ⢠14
Token Bottleneck: One Token to Remember Dynamics Paper ⢠2507.06543 ⢠Published Jul 9, 2025 ⢠20
HyperCLOVA X SEED Collection HyperCLOVA X SEED is NAVER's lightweight open-source lineup with a strong focus on Korean language performance ⢠6 items ⢠Updated Dec 24, 2025 ⢠43
ProLIP Collection Official ProLIP weights, Probabilistic Language-Image Pre-Training (ICLR 2025) ⢠7 items ⢠Updated Apr 18, 2025 ⢠10
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation Paper ⢠2411.19067 ⢠Published Nov 28, 2024 ⢠8
Cosmos-Tokenizer1 Collection ā ļø This collection is archived. š https://huggingface.co/collections/nvidia/cosmos3 ⢠22 items ⢠Updated Aug 11 ⢠44