Reason Through the Latent! Making Latent Visual Reasoning Necessary Paper • 2609.06746 • Published 21 days ago • 36
nota-ai/Nemotron-3.5-Lightning-30B-A3B-NVFP4-Global-Pruned-15 Text Generation • 16B • Updated Aug 17 • 190 • 21
Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models Paper • 2512.10362 • Published Dec 11, 2025 • 2
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval Paper • 2607.04605 • Published Jul 6 • 22
Running 243 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 243 Building and scaling RL environments for LLM training
QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering Paper • 2604.24052 • Published Apr 27
view article Article Introducing Storage Buckets on the Hugging Face Hub +10 Wauplin, coyotte508, XciD, victor, julien-c, lhoestq, pierric, Sylvestre, hlarcher, rajatarya, seanses, assafvayner • Mar 10 • 198