-
DevQuasar/nvidia.Llama-3_1-Nemotron-Ultra-253B-CPT-v1-GGUF
Text Generation • 253B • Updated • 73 -
DevQuasar/meta-llama.Llama-4-Maverick-17B-128E-Instruct-GGUF
Text Generation • 401B • Updated • 222 -
DevQuasar/meta-llama.Llama-4-Scout-17B-16E-Instruct-GGUF
Text Generation • 108B • Updated • 165 • 3 -
DevQuasar/nvidia.Llama-3_1-Nemotron-Ultra-253B-v1-GGUF
Text Generation • 253B • Updated • 155 • 7
Collections
Discover the best community collections!
Collections trending this week
-
meta-llama/Llama-3.2-11B-Vision
Image-Text-to-Text • 11B • Updated • 8.93k • 595 -
meta-llama/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 110k • 1.63k -
meta-llama/Llama-3.2-90B-Vision
Image-Text-to-Text • 89B • Updated • 114 • 134 -
meta-llama/Llama-3.2-90B-Vision-Instruct
Image-Text-to-Text • 89B • Updated • 125k • 359
-
ggml-org/Qwen2.5-Coder-0.5B-Q8_0-GGUF
Text Generation • 0.5B • Updated • 491 • 10 -
ggml-org/Qwen2.5-Coder-1.5B-Q8_0-GGUF
Text Generation • 2B • Updated • 4.37k • 19 -
ggml-org/Qwen2.5-Coder-3B-Q8_0-GGUF
Text Generation • 3B • Updated • 1.74k • 11 -
ggml-org/Qwen2.5-Coder-7B-Q8_0-GGUF
Text Generation • 8B • Updated • 2.14k • 10
-
moonshotai/Kimi-VL-A3B-Thinking-2506
Image-Text-to-Text • 16B • Updated • 8.5k • 373 -
moonshotai/Kimi-VL-A3B-Instruct
Image-Text-to-Text • 16B • Updated • 388k • 277 -
moonshotai/Kimi-VL-A3B-Thinking
Image-Text-to-Text • 16B • Updated • 171k • 450 -
moonshotai/MoonViT-SO-400M
Image Feature Extraction • 0.4B • Updated • 3.84k • 90
-
MaziyarPanahi/Meta-Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 87k • 37 -
mradermacher/DeepSeek-R1-Distill-Llama-8B-Abliterated-GGUF
8B • Updated • 3.49k • 40 -
mlabonne/Meta-Llama-3.1-8B-Instruct-abliterated-GGUF
8B • Updated • 34.4k • 201 -
bartowski/Meta-Llama-3.1-8B-Instruct-abliterated-GGUF
Text Generation • 8B • Updated • 8.47k • 14
-
DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior
Paper • 2310.16818 • Published • 33 -
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Paper • 2401.02954 • Published • 56 -
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Paper • 2401.06066 • Published • 63 -
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Paper • 2401.14196 • Published • 74
-
Qwen2.5 Coder Artifacts
🐢1.73kGenerate and preview app code from a text description
-
Qwen/Qwen2.5-Coder-32B-Instruct
Text Generation • 33B • Updated • 1.18M • • 2.09k -
Qwen/Qwen2.5-Coder-32B
Text Generation • 33B • Updated • 2.34k • • 159 -
Qwen2.5-Coder Technical Report
Paper • 2409.12186 • Published • 158
-
DevQuasar/nvidia.Llama-3_1-Nemotron-Ultra-253B-CPT-v1-GGUF
Text Generation • 253B • Updated • 73 -
DevQuasar/meta-llama.Llama-4-Maverick-17B-128E-Instruct-GGUF
Text Generation • 401B • Updated • 222 -
DevQuasar/meta-llama.Llama-4-Scout-17B-16E-Instruct-GGUF
Text Generation • 108B • Updated • 165 • 3 -
DevQuasar/nvidia.Llama-3_1-Nemotron-Ultra-253B-v1-GGUF
Text Generation • 253B • Updated • 155 • 7
-
moonshotai/Kimi-VL-A3B-Thinking-2506
Image-Text-to-Text • 16B • Updated • 8.5k • 373 -
moonshotai/Kimi-VL-A3B-Instruct
Image-Text-to-Text • 16B • Updated • 388k • 277 -
moonshotai/Kimi-VL-A3B-Thinking
Image-Text-to-Text • 16B • Updated • 171k • 450 -
moonshotai/MoonViT-SO-400M
Image Feature Extraction • 0.4B • Updated • 3.84k • 90
-
MaziyarPanahi/Meta-Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 87k • 37 -
mradermacher/DeepSeek-R1-Distill-Llama-8B-Abliterated-GGUF
8B • Updated • 3.49k • 40 -
mlabonne/Meta-Llama-3.1-8B-Instruct-abliterated-GGUF
8B • Updated • 34.4k • 201 -
bartowski/Meta-Llama-3.1-8B-Instruct-abliterated-GGUF
Text Generation • 8B • Updated • 8.47k • 14
-
meta-llama/Llama-3.2-11B-Vision
Image-Text-to-Text • 11B • Updated • 8.93k • 595 -
meta-llama/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 110k • 1.63k -
meta-llama/Llama-3.2-90B-Vision
Image-Text-to-Text • 89B • Updated • 114 • 134 -
meta-llama/Llama-3.2-90B-Vision-Instruct
Image-Text-to-Text • 89B • Updated • 125k • 359
-
DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior
Paper • 2310.16818 • Published • 33 -
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Paper • 2401.02954 • Published • 56 -
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Paper • 2401.06066 • Published • 63 -
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Paper • 2401.14196 • Published • 74
-
ggml-org/Qwen2.5-Coder-0.5B-Q8_0-GGUF
Text Generation • 0.5B • Updated • 491 • 10 -
ggml-org/Qwen2.5-Coder-1.5B-Q8_0-GGUF
Text Generation • 2B • Updated • 4.37k • 19 -
ggml-org/Qwen2.5-Coder-3B-Q8_0-GGUF
Text Generation • 3B • Updated • 1.74k • 11 -
ggml-org/Qwen2.5-Coder-7B-Q8_0-GGUF
Text Generation • 8B • Updated • 2.14k • 10
-
Qwen2.5 Coder Artifacts
🐢1.73kGenerate and preview app code from a text description
-
Qwen/Qwen2.5-Coder-32B-Instruct
Text Generation • 33B • Updated • 1.18M • • 2.09k -
Qwen/Qwen2.5-Coder-32B
Text Generation • 33B • Updated • 2.34k • • 159 -
Qwen2.5-Coder Technical Report
Paper • 2409.12186 • Published • 158