MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 360 • 167 litert-community/gemma-4-E2B-it-litert-lm Updated 22 days ago • 1.23M • 436 Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 360 • 167 litert-community/gemma-4-E2B-it-litert-lm Updated 22 days ago • 1.23M • 436 Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
Mer0vin8ian/moonshine-streaming-small-onnx Automatic Speech Recognition • Updated Jul 13 • 6 • 1