UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering Paper • 2605.30076 • Published May 28 • 26
Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch Paper • 2311.03099 • Published Nov 6, 2023 • 38
Closed-Form Spectral Regularization for Multi-Task Model Merging Paper • 2606.07289 • Published Jun 5 • 2
HiPO: Hybrid Policy Optimization for Dynamic Reasoning in LLMs Paper • 2509.23967 • Published Sep 28, 2025 • 5
Nemotron-Personas Collection A collection of multilingual, region-specific synthetic persona datasets that support sovereign AI development across many countries and regions. • 10 items • Updated Aug 11 • 76
Nemotron-Post-Training-v3 Collection Collection of datasets used in the post-training phase of Nemotron Nano, Super, and Ultra v3. • 50 items • Updated Aug 11 • 203
Nemotron Agentic & Tool-Use Collection Datasets for building models capable of function calling, multi-step agentic tasks, terminal use, and SWE workflows. • 11 items • Updated Aug 11 • 24
Nemotron-Pre-Training-Datasets Collection Large scale pre-training datasets used in the Nemotron family of models. • 15 items • Updated Aug 11 • 195
Nemotron Reward Modeling Collection Human preference data, reward model training sets, and generative reward modeling data for training Nemotron reward models. • 6 items • Updated Aug 11 • 5
Nemotron Math & Reasoning Collection Datasets for building models that excel at math reasoning, proofs, and quantitative problem-solving. Covers SFT, RL, and pretraining data. • 23 items • Updated Aug 11 • 16
Nemotron Chat & Instruction Following Collection Datasets for building helpful, multi-turn, instruction-following conversational models across single and multi-turn settings. • 19 items • Updated Aug 11 • 10
Nemotron Code & SWE Collection Datasets for building models that write, debug, and reason about code. Covers competitive programming, software engineering, and code pretraining. • 14 items • Updated Aug 11 • 9
Nemotron Supervised Fine-Tuning Collection SFT datasets covering math, code, chat, safety, agentic, VLM, multilingual, and specialized domains. • 44 items • Updated Aug 11 • 22