bf16 layer is no good for Vulkan.. had to quant it to Q8 to make this model usable on my dual v620s

#5
by Black6spdZ - opened

bf16 layer is no good for Vulkan.. had to quant it to Q8 to make this model usable on my dual v620s

That sounds a whole lot like the vulkan maxMemoryAllocationSize limit mentioned in another thread https://huggingface.co/DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF/discussions/39

oh boy.. can't wait to see this voodoo applied to the new qwen3.8-27b!

Sign up or log in to comment