anyway to make this a tad smaller?

#5
by Dragonite9000 - opened

Just curious if its possible, if there is a layer we could drop. I would totally forego vision completely if it made this half a gb smaller.

Get byteshape/Qwen3.8-27B-GGUF its smaller, but also you dont need to load vision if you dont want, use llama.cpp directly and you dont need to pass the mmproj file for vision, saves around 1gb

Sign up or log in to comment