Je Lee
geveent
·
AI & ML interests
None yet
Recent Activity
liked a model 6 days ago
rogerai-fyi/DeepSeek-V4-Flash-MTP-GGUF reacted to danielhanchen's post with 🔥 10 days ago
We’re releasing new Qwen3.6 quants that run 2.5× faster on your GPU. ⚡
Qwen3.6-27B NVFP4 runs on 24GB VRAM.
35B-A3B can hit 17,561 tok/s (B200).
We also improved accuracy, tool calling, agent use, and looping.
Qwen3.6 NVFP4: https://huggingface.co/collections/unsloth/nvfp4
Guide: https://unsloth.ai/docs/models/qwen3.6#nvfp4 new activity 13 days ago
antirez/deepseek-v4-gguf:Running with recently merged llama.cpp PROrganizations
None yet