Qwen3.5 4B Q4_K_M GGUF

This repository contains the exact immutable local-model asset used by NovelAide.

Provenance

The base model is Qwen/Qwen3.5-4B; NovelAide consumes the LM Studio community GGUF quantization lmstudio-community/Qwen3.5-4B-GGUF.

Files

The NovelAide local package includes an upstream README.md. Hugging Face reserves that path for this model card, so the original bytes are published as UPSTREAM_README.md; its size and SHA-256 below remain unchanged.

File Size SHA-256
UPSTREAM_README.md 0.00 MiB e1b3ca9c804ec81f9d0f8e51eab46d4e0e1b18521e7436bcb95736fb3d0414ae
Qwen3.5-4B-Q4_K_M.gguf 2582.09 MiB 25082a7dd3776cc3c741c6347d3bd04523f05796607b3fbc32fa3a25dfa1418c

Usage

Use the GGUF files with a compatible llama.cpp runtime and the ONNX files with a compatible ONNX / Transformers.js runtime. NovelAide pins the exact files and checksums shown above; do not substitute similarly named quantizations.

License and attribution

The model asset follows the upstream apache-2.0 license. Review the linked base model and asset source model cards for their complete terms, limitations, and attribution requirements. NovelAide is not affiliated with or endorsed by the upstream model authors.

Downloads last month
115
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for novelaide/Qwen3.5-4B-Q4_K_M-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(405)
this model