MiniMax-H3 Collection NVFP4 text encoder for MiniMax-H3 video generation: the Qwen3-VL encoder quantized to run in ComfyUI on a single card. • 1 item • Updated 6 days ago
Qwen3.8 Collection Qwen3.8-27B (dense, vision) imatrix GGUF: 18 tiers plus q8_0/f16 mmproj and an MTP draft. More Qwen3.8 sizes as they land. • 1 item • Updated 6 days ago
DeepSeek-V4 Collection Sub-4-bit GGUF quants of DeepSeek-V4: Flash-0731 (seven PPL-tested tiers + DSpark draft) and Pro-0813 (1.57T, factory FP4 master). • 2 items • Updated 7 days ago
DeepSeek-V4 Collection Sub-4-bit GGUF quants of DeepSeek-V4: Flash-0731 (seven PPL-tested tiers + DSpark draft) and Pro-0813 (1.57T, factory FP4 master). • 2 items • Updated 7 days ago
Kimi-K3 Collection Moonshot Kimi-K3 GGUF quants, imatrix-calibrated. For llama.cpp / Ollama / LM Studio. • 1 item • Updated 22 days ago
Qwen3 Collection All Qwen3 quantized builds: GGUF, FP8, AWQ and GPTQ. Covers 4B to 32B dense and 30B-A3B MoE. • 20 items • Updated 23 days ago