MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 446 • 165 litert-community/gemma-4-E2B-it-litert-lm Updated 11 days ago • 1.06M • 402 Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
MLLM NexaAI/OmniNeural-4B Any-to-Any • Updated Nov 7, 2025 • 446 • 165 litert-community/gemma-4-E2B-it-litert-lm Updated 11 days ago • 1.06M • 402 Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
Running 14 TurboQuant on Consumer GPUs — 100K Context on RTX 3090, 64K on RTX 4070 🚀 14 Extend LLM context to 100K tokens on consumer GPUs
Mer0vin8ian/moonshine-streaming-small-onnx Automatic Speech Recognition • Updated Jul 13 • 11 • 1