Inference Providers
Active filters: mamba2
AEON-7/Nemotron-3-Nano-Omni-AEON-Ultimate-Uncensored-BF16
Any-to-Any
• 33B • Updated • 305
• 6
S4MPL3BI4S/Nemotron-3-Nano-4B-Coding-Agent-GGUF
Text Generation
• 4B • Updated • 2.36k
• 5
VAREQON/Nemotron-3-Nano-4B-RotorQuant-GGUF-Q2_K
Text Generation
• 4B • Updated • 119
• 1
Premshay/Nemotron-3-Super-120B-A12B-MTP-GGUF
Text Generation
• 124B • Updated • 494
• 1
MRockatansky/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-MTP-GGUF
Text Generation
• 78B • Updated • 1
kuzyn8/Nemotron-3-Nano-4B-RotorQuant-GGUF-Q2_K
Text Generation
• 4B • Updated • 123
• 2
0.1B • Updated • 88
• 2
0.4B • Updated • 12
• 1
0.8B • Updated • 11
1B • Updated • 14
3B • Updated • 11
• 1
0.1B • Updated • 10.8k
• 7
0.4B • Updated • 3.33k
• 2
0.8B • Updated • 318
• 2
1B • Updated • 945
• 1
3B • Updated • 2.99k
• 3
27.7M • Updated • 2.07k
8B • Updated • 265
• 1
mradermacher/SphinX-i1-GGUF
8B • Updated • 477
• 1
Text Generation
• 20.3M • Updated • 52
• 1
petkopetkov/mamba2-130m-hf
0.2B • Updated • 55
petkopetkov/mamba2-370m-hf
0.4B • Updated • 8
petkopetkov/mamba2-780m-hf
0.9B • Updated • 12
• 1
petkopetkov/mamba2-1.3b-hf
1B • Updated • 12
petkopetkov/mamba2-2.7b-hf
3B • Updated • 135
weathermanj/NVIDIA-Nemotron-Nano-9B-v2-gguf
Text Generation
• 9B • Updated • 242
• 1
mlx-community/granite-4.0-h-tiny-3bit-MLX
Text Generation
• 0.9B • Updated • 71
• 2
mlx-community/granite-4.0-h-tiny-6bit-MLX
Text Generation
• 7B • Updated • 46
• 1
mlx-community/granite-4.0-h-tiny-5bit-MLX
Text Generation
• 1B • Updated • 21
• 2
ibm-research/materials.str-bamba
Updated • 8
• 3