MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 622 • 167 litert-community/gemma-4-E2B-it-litert-lm Updated 2 days ago • 1.04M • 413 Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 622 • 167 litert-community/gemma-4-E2B-it-litert-lm Updated 2 days ago • 1.04M • 413 Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
Running 15 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 15 Extend LLM context to 100K tokens on consumer GPUs
Mer0vin8ian/moonshine-streaming-small-onnx Automatic Speech Recognition • Updated Jul 13 • 6 • 1