Collections
Discover the best community collections!
Collections trending this week
-
deepreinforce-ai/Ornith-1.0-9B-GGUF
Text Generation • 9B • Updated • 4.07M • 576 -
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
Image-Text-to-Text • 9B • Updated • 90k • 286 -
DJLougen/Qwable-5-27B-Coder
Text Generation • 28B • Updated • 359 • • 42 -
DJLougen/Qwable-5-27B-Coder-GGUF
Text Generation • 27B • Updated • 9.43k • 16
-
Attention Is All You Need
Paper • 1706.03762 • Published • 134 -
FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
Paper • 2205.14135 • Published • 15 -
Efficient Memory Management for Large Language Model Serving with PagedAttention
Paper • 2309.06180 • Published • 64 -
FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Paper • 2307.08691 • Published • 9
-
YoozLabs/gemma-4-E4B-it-qat-lean-4bit-mlx
Text Generation • 1B • Updated • 192 -
YoozLabs/gemma-4-12B-it-qat-lean-4bit-mlx
Text Generation • 12B • Updated • 269 -
YoozLabs/Qwen3.5-0.8B-qat-lean-4bit-mlx
Text Generation • 0.1B • Updated • 129 -
YoozLabs/Qwen3.5-4B-qat-lean-4bit-mlx
Text Generation • 0.7B • Updated • 171
-
Attention Is All You Need
Paper • 1706.03762 • Published • 134 -
FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
Paper • 2205.14135 • Published • 15 -
Efficient Memory Management for Large Language Model Serving with PagedAttention
Paper • 2309.06180 • Published • 64 -
FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Paper • 2307.08691 • Published • 9
-
YoozLabs/gemma-4-E4B-it-qat-lean-4bit-mlx
Text Generation • 1B • Updated • 192 -
YoozLabs/gemma-4-12B-it-qat-lean-4bit-mlx
Text Generation • 12B • Updated • 269 -
YoozLabs/Qwen3.5-0.8B-qat-lean-4bit-mlx
Text Generation • 0.1B • Updated • 129 -
YoozLabs/Qwen3.5-4B-qat-lean-4bit-mlx
Text Generation • 0.7B • Updated • 171
-
deepreinforce-ai/Ornith-1.0-9B-GGUF
Text Generation • 9B • Updated • 4.07M • 576 -
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
Image-Text-to-Text • 9B • Updated • 90k • 286 -
DJLougen/Qwable-5-27B-Coder
Text Generation • 28B • Updated • 359 • • 42 -
DJLougen/Qwable-5-27B-Coder-GGUF
Text Generation • 27B • Updated • 9.43k • 16