Josh Leverette
coder543
AI & ML interests
LLMs
Recent Activity
new activity 5 days ago
inclusionAI/LLaDA2.2-mini:vLLM or llama.cpp support? updated a model 5 days ago
coder543/LLaDA2.2-flash-minnow updated a model 5 days ago
coder543/LLaDA2.2-mini-minnowOrganizations
None yet
vLLM or llama.cpp support?
1
#1 opened 6 days ago
by
coder543
Tested on the Artificial Analysis Intelligence Index?
👀 2
3
#25 opened about 1 month ago
by
ATHENIS
Why is it much slower and much more memory hungry than slightly larger models such as Qwen3.5-4B?
6
#31 opened about 1 month ago
by
slightlyoutofphase
GGUF
🚀🔥 4
12
#5 opened 3 months ago
by
wepiqx
QAD Q4_K_M?
👀 1
#8 opened 25 days ago
by
coder543
MTP support?
➕ 13
2
#11 opened 26 days ago
by
coder543
MTP seems to be missing?
2
#1 opened 27 days ago
by
coder543
New GGUFs required? "chat : add new template for DeepSeek V4 Flash 0731 (#26398)"
👍 4
5
#28 opened about 1 month ago
by
rtzurtz
Thanks for the release!
❤️ 7
1
#1 opened about 2 months ago
by
coder543
1-bit Kimi K3 vs Claude Opus 5 vs GPT 5.6
👀🚀 20
24
#12 opened about 2 months ago
by
danielhanchen
Upstream the `laguna` branch?
➕ 1
2
#19 opened about 2 months ago
by
coder543
DFlash configuration issue causing acceptance rate collapse
👍 7
2
#6 opened about 2 months ago
by
coder543
LFM2.5-230M beats LFM2.5-350M on GSM8K?
👀 1
2
#10 opened 3 months ago
by
coder543
Chat template that supports reasoning
5
#1 opened 3 months ago
by
coder543
Chat template issues with multiple rounds of tool calling
➕ 2
11
#115 opened 3 months ago
by
Kimahriman
GGUF advertises wrong max context?
6
#2 opened 3 months ago
by
coder543
Context length is at 131k instead of 256k
5
#8 opened 3 months ago
by
Orangeswim
Wrong context size in config.json?
👍 1
1
#10 opened 3 months ago
by
coder543
Benchmarks of reasoning levels?
1
#6 opened 4 months ago
by
coder543
correct max context length to 200k
1
#1 opened 4 months ago
by
walterbm-cohere