Rio Grande
dandandelion
AI & ML interests
None yet
Recent Activity
new activity 2 days ago
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF:Just appreciation, nothing to see here new activity 4 days ago
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF:Anyone unable to load the model on LLama.cpp? new activity 4 days ago
pfeifferj/Qwen3.8-Flash-Next-GSQ-RCO-GGUF:It's not perfectOrganizations
None yet
Just appreciation, nothing to see here
5
#13 opened 3 days ago
by
dandandelion
Anyone unable to load the model on LLama.cpp?
5
#38 opened 5 days ago
by
Iloveprogramming342
It's not perfect
👍 1
3
#3 opened 5 days ago
by
dandandelion
Run UD-Q4-K-XL on 16GB VRAM and 32 GB RAM at 15+ tok/s
👀 1
10
#65 opened 14 days ago
by
Apolog1ze-Dev
Normal speed for my setup (or under-performing)?
8
#3 opened 13 days ago
by
dandandelion
Incredible speed and ability
3
#16 opened 19 days ago
by
jdluzen
How to keep n-gram table on fast nvme ssd
20
#23 opened 25 days ago
by
mayankiit04
Request for a Faster UD-Q4_K_XL Quantization of Qwen3.8-Flash-Next
2
#14 opened 19 days ago
by
NamerPRO
NGram to SSD streaming?
👀 1
8
#1 opened 25 days ago
by
dandandelion
Suggestion: Support for Ngram SSD Offloading in Unsloth Desktop
➕ 28
4
#15 opened 25 days ago
by
wuyule
Perplexity comparison: Ridge 3.7bpw vs UD-Q3_K_XL
❤️ 3
4
#5 opened about 1 month ago
by
enricos
Excessive CoT verbosity on simple, well-defined tasks
9
#5 opened about 1 month ago
by
desugar
M quants??
5
#1 opened about 1 month ago
by
pawarshardul
Squeezing Qwen 3.8 27B into a Single 16 GB GPU — Almost 42 tok/s, 64K Context
🚀 2
7
#70 opened about 1 month ago
by
Ataa
Q3 K M (13.8 GB) is larger than UD Q3 K XL(13.4GB)
👍 1
3
#58 opened about 1 month ago
by
whatitall
Using this on vllm but excessive tbinking
👀➕ 3
8
#13 opened about 1 month ago
by
crystech
UD quants for 16GB VRAM?
7
#35 opened about 1 month ago
by
Witherhoard
UD_IQ1_XXXS possible like you did with the big one?
🔥 1
4
#60 opened about 1 month ago
by
TheWegemann
Not great, not horrible
🚀 4
2
#14 opened about 1 month ago
by
tandaraSa