Inference Providers
Active filters: mamba2
AEON-7/Nemotron-3-Nano-Omni-AEON-Ultimate-Uncensored-BF16
Any-to-Any
• 33B • Updated • 283
• 6
S4MPL3BI4S/Nemotron-3-Nano-4B-Coding-Agent-GGUF
Text Generation
• 4B • Updated • 2.38k
• 5
VAREQON/Nemotron-3-Nano-4B-RotorQuant-GGUF-Q2_K
Text Generation
• 4B • Updated • 124
• 1
Premshay/Nemotron-3-Super-120B-A12B-MTP-GGUF
Text Generation
• 124B • Updated • 516
• 1
MRockatansky/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-MTP-GGUF
Text Generation
• 78B • Updated • 719
• 1
kuzyn8/Nemotron-3-Nano-4B-RotorQuant-GGUF-Q2_K
Text Generation
• 4B • Updated • 133
• 2
0.1B • Updated • 88
• 2
0.4B • Updated • 11
• 1
0.8B • Updated • 12
1B • Updated • 14
3B • Updated • 11
• 1
0.1B • Updated • 9.17k
• 7
0.4B • Updated • 3k
• 2
0.8B • Updated • 306
• 2
1B • Updated • 907
• 1
3B • Updated • 2.85k
• 3
27.7M • Updated • 2.07k
8B • Updated • 265
• 1
mradermacher/SphinX-i1-GGUF
8B • Updated • 494
• 1
Text Generation
• 20.3M • Updated • 53
• 1
petkopetkov/mamba2-130m-hf
0.2B • Updated • 57
petkopetkov/mamba2-370m-hf
0.4B • Updated • 11
petkopetkov/mamba2-780m-hf
0.9B • Updated • 14
• 1
petkopetkov/mamba2-1.3b-hf
1B • Updated • 13
petkopetkov/mamba2-2.7b-hf
3B • Updated • 142
weathermanj/NVIDIA-Nemotron-Nano-9B-v2-gguf
Text Generation
• 9B • Updated • 253
• 1
mlx-community/granite-4.0-h-tiny-3bit-MLX
Text Generation
• 0.9B • Updated • 74
• 2
mlx-community/granite-4.0-h-tiny-6bit-MLX
Text Generation
• 7B • Updated • 48
• 1
mlx-community/granite-4.0-h-tiny-5bit-MLX
Text Generation
• 1B • Updated • 22
• 2
ibm-research/materials.str-bamba
Updated • 10
• 3