MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 117 • 165 litert-community/gemma-4-E2B-it-litert-lm Updated 28 days ago • 1.09M • 387 Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 117 • 165 litert-community/gemma-4-E2B-it-litert-lm Updated 28 days ago • 1.09M • 387 Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
Mer0vin8ian/moonshine-streaming-small-onnx Automatic Speech Recognition • Updated 26 days ago • 22 • 1