RuxAI DeepSeek-R1 Distill Qwen3.5 4B GGUF

A fine-tuned GGUF version of Qwen3.5-4B optimized for reasoning tasks.

This model was created by fine-tuning Qwen3.5-4B on:

  • Dataset: a-m-team/AM-DeepSeek-R1-Distilled-1.4M
  • Training style: DeepSeek-R1 style reasoning distillation
  • Format: GGUF for llama.cpp and compatible runtimes

Model Details

  • Base model: Qwen3.5-4B
  • Architecture: Qwen3.5
  • Quantization: GGUF
  • Supported runtimes:
    • llama.cpp
    • LM Studio
    • Ollama (with conversion)
    • other GGUF-compatible tools

Usage

Example with llama.cpp:

llama-cli \
  -m RuxAI-deepseekr1-distill-qwen3.5-4B-Q8_0.gguf \
  -p "Explain why the sky is blue."
Downloads last month
1,641
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for PavelH-cz/RuxAI-deepseekr1-distill-qwen3.5-4B-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(9)
this model

Dataset used to train PavelH-cz/RuxAI-deepseekr1-distill-qwen3.5-4B-GGUF