KeyboardMasher/Step-3.5-Flash-MTP-GGUF-backup
518 GB
Why Qwen 3.5 and not 3.6?
At 4 bit or lower use IQ-type quant. The math is more advanced and quantization error is lower. You can double the context with -ctk q8_0 -ctv q8_0for virtually no loss of quality and speed.