no support llama.cpp b11404

#1
by astromc - opened

no support llama.cpp b11404 ver: Windows x64 (Vulkan) and Windows x64 (CPU)
llama cli -hf Cactus-Compute/gemma-4-e2b-it-hybrid-GGUF:Q4_K_M

Downloading mmproj-F16.gguf ──────────────────────────────────────── 100%
Downloading gemma-4-e2b-it-hybrid-Q4_K_M.gguf ────────────────────── 100%

Loading model... \25.26.957.470 E llama_model_load: error loading model: unknown model architecture: 'gemma-4-e2b-it-hybrid'
25.26.957.477 E llama_model_load_from_file_impl: failed to load model
25.26.957.513 E common_fit_params: encountered an error while trying to fit params to free device memory: failed to load model |25.27.429.751 E llama_model_load: error loading model: unknown model architecture: 'gemma-4-e2b-it-hybrid'
25.27.429.756 E llama_server exited with code 1
llama_model_load_from_file_impl: failed to load model
25.27.429.762 E cmn common_init_: failed to load model '.cache\huggingface\hub\models--Cactus-Compute--gemma-4-e2b-it-hybrid-GGUF\snapshots\4786442139d63b0962728e037470adf0eaca2988\gemma-4-e2b-it-hybrid-Q4_K_M.gguf'
25.27.429.766 E srv load_model: failed to load model, '.cache\huggingface\hub\models--Cactus-Compute--gemma-4-e2b-it-hybrid-GGUF\snapshots\4786442139d63b0962728e037470adf0eaca2988\gemma-4-e2b-it-hybrid-Q4_K_M.gguf'
25.27.430.587 E srv llama_server: exiting due to model loading error Error: the server exited before becoming ready

Ok, Resolved.
a custom one is required llama.cpp - https://github.com/cactus-compute/cactus-hybrid

Sign up or log in to comment