view article Article Welcome Inkling by Thinking Machines +3 burtenshaw, merve, pcuenq, ariG23498, andito ⢠Jul 15 ⢠167
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver Paper ⢠2604.08377 ⢠Published Apr 9 ⢠226
view article Article NVIDIA Cosmos Reason 2 Brings Advanced Reasoning To Physical AI nvidia ⢠Jan 5 ⢠65
view article Article We Got Claude to Fine-Tune an Open Source LLM burtenshaw, evalstate ⢠Dec 4, 2025 ⢠634
SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models Paper ⢠2504.11468 ⢠Published Apr 10, 2025 ⢠30
SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training Paper ⢠2501.17161 ⢠Published Jan 28, 2025 ⢠128
Cosmos-Predict2 Collection ā ļø This collection is archived. š https://huggingface.co/collections/nvidia/cosmos-predict25 ⢠10 items ⢠Updated Aug 11 ⢠40
MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers Paper ⢠2508.20453 ⢠Published Aug 28, 2025 ⢠63
view article Article Small Language Models (SLM): A Comprehensive Overview jjokah ⢠Feb 22, 2025 ⢠177
Wan: Open and Advanced Large-Scale Video Generative Models Paper ⢠2503.20314 ⢠Published Mar 26, 2025 ⢠72
Physical AI Collection Collection of open, commercial-grade datasets for physical AI developers ⢠57 items ⢠Updated Aug 11 ⢠182
AceReason Collection Math and Code reasoning model trained through reinforcement learning (RL) ⢠7 items ⢠Updated Aug 11 ⢠22
Reward Models 06-2025 Collection Nemotron reward models. For use in RLHF pipelines and LLM-as-a-Judge ⢠8 items ⢠Updated Aug 11 ⢠25
Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms Paper ⢠2310.07161 ⢠Published Oct 11, 2023 ⢠1
Qwen2.5-1M Collection The long-context version of Qwen2.5, supporting 1M-token context lengths ⢠3 items ⢠Updated Dec 31, 2025 ⢠128
OpenReasoning-Nemotron Collection Collection of models for OpenReasoning-Nemotron which are trained on 5M reasoning traces for Math, Code and Science. ⢠6 items ⢠Updated Aug 11 ⢠48