DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 6 days ago • 160
On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Paper • 2608.30320 • Published 23 days ago • 63
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper • 2609.00111 • Published 23 days ago • 313
Cosmos-Reason2 Collection ⚠️ This collection is archived. 👉 https://huggingface.co/collections/nvidia/cosmos3 • 8 items • Updated Aug 11 • 28
view article Article Welcome NVIDIA Cosmos 3: The First Open Omni-model for Physical AI Reasoning and Action nvidia • Jun 1 • 90
NVIDIA OmniDreams Collection NVIDIA OmniDreams model checkpoints and sample datasets. • 3 items • Updated Aug 11 • 10
view article Article Ulysses Sequence Parallelism: Training with Million-Token Contexts kashif, stas • Mar 9 • 33
view article Article NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics nvidia • Jul 27 • 77
Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Paper • 2505.23705 • Published May 29, 2025 • 1
LingoQA: Video Question Answering for Autonomous Driving Paper • 2312.14115 • Published Dec 21, 2023 • 4
Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published Jul 10 • 83
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning Paper • 2607.07508 • Published Jul 8 • 33