phak
phakio
AI & ML interests
None yet
Recent Activity
new activity 15 days ago
unsloth/DeepSeek-V4-Flash-0731-GGUF:Great agentic results even with IQ1_S! new activity 19 days ago
sokann/DeepSeek-V4-Flash-0731-GGUF:Solid results in ik_llama.cpp! liked a model 19 days ago
sokann/DeepSeek-V4-Flash-0731-GGUFOrganizations
None yet
Great agentic results even with IQ1_S!
👍🤗 8
5
#19 opened 18 days ago
by
phakio
Solid results in ik_llama.cpp!
🔥 2
#1 opened 19 days ago
by
phakio
Working Results + Quick Benchmark
❤️ 10
7
#1 opened 23 days ago
by
pwdz
Pretty solid results with ik_llama and OpenCode!
🔥 1
11
#2 opened 2 months ago
by
phakio
Running with recently merged llama.cpp PR
👍 6
6
#16 opened about 2 months ago
by
ubergarm
Issues with UD-Q2_K_XL quant
5
#9 opened about 2 months ago
by
labhraighlep
Q2_K has endless thinking loop issue
2
#2 opened about 2 months ago
by
phakio
really awesome speeds! running at 256k context.
🔥 1
6
#11 opened 4 months ago
by
mtcl
GLM-5.2 GGUF Benchmarks!
❤️🔥 14
26
#3 opened 2 months ago
by
danielhanchen
Thanks for the quick quants! Reccomended mmproj?
2
#1 opened 2 months ago
by
phakio
Any plans for reviving this model with MTP support?
👍 1
3
#14 opened 3 months ago
by
phakio
Thanks for the quantizations, can we get MTP Qwen 3.5 397B GGUF?
2
#5 opened 3 months ago
by
tidjei43
Running good on full GPU offload (1x4090, 3x3090) (Multi-GPU Offload Crash Fix)
🔥 1
#2 opened 3 months ago
by
phakio
The model is working okay! (Temporary fix for stop token being ignored)
5
#1 opened 4 months ago
by
phakio
How to use MTP in GGUF?
20
#2 opened 4 months ago
by
Friedland
Running great on my Intel QYFS and DDR5 only! (CUDA gives error)
2
#1 opened 4 months ago
by
phakio
Working good on 96GB VRAM + DDR5 Setup
❤️ 1
5
#2 opened 4 months ago
by
phakio
Great model for single GPU use cases.
🔥 4
16
#1 opened 4 months ago
by
phakio
A fun, quick little model!
#1 opened 5 months ago
by
phakio
Performance on Intel QYFS, 512GB DDR5 and 96GB VRAM
👍 1
6
#3 opened 7 months ago
by
phakio