Instructions to use midium-ai/decider-4b-dwq-4bit with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use midium-ai/decider-4b-dwq-4bit with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] hf download midium-ai/decider-4b-dwq-4bit --local-dir decider-4b-dwq-4bit
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
Download decider_config.json from midium-ai/decider-4b-dwq-4bit: direct link, hf CLI and curl.
- Browser
- Download file 1.38 kB
-
https://huggingface.co/midium-ai/decider-4b-dwq-4bit/resolve/main/decider_config.json
- Command line
-
hf download hf://midium-ai/decider-4b-dwq-4bit/decider_config.json
-
curl -L -o decider_config.json https://huggingface.co/midium-ai/decider-4b-dwq-4bit/resolve/main/decider_config.json
1.38 kB
| { | |
| "temperature": 1.099, | |
| "temperature_by_type": { | |
| "choice": 1.11, | |
| "noul": 1.56, | |
| "score": 1.287 | |
| }, | |
| "neutralize_none": false, | |
| "version": "4b-v2.1", | |
| "base": "Mapika/decider-4b v1 + LoRA (merged); v1 is Qwen/Qwen3.5-4B-Base + one supervised pass over mixture v2", | |
| "layout": "plain", | |
| "max_options": 255, | |
| "max_state_tokens": 32768, | |
| "schema_first": false, | |
| "schema_first_trained": false, | |
| "isolated_levels": true, | |
| "release_date": "2026-09-24", | |
| "requires": "decider-ai>=1.4.0 for temperature_by_type; older versions serve every answer at temperature", | |
| "stage": "decider-4b v1 + LoRA rank 64 (alpha 128) on attention and MLP, LR 1e-4, 2 epochs (1,518 steps of 65,536 tokens) over v2's 29,325-row mix in the plain state-first layout (generated decision families with code-computed answers, Qwen3.6-27B-written document questions kept when two independent answers agreed, human-labelled public sets, replay of v1's mixture v2), with the replay rows trained toward v1's own answer distribution (KL to v1) instead of their labels, merged into the bf16 weights; no RL stage; temperature fitted by NLL on 61 in-task regression tasks (the 67 in-task tasks without banking77, clinc_oos, mmlu, arc, winogrande, hellaswag); temperature_by_type fitted with decider.calibrate.fit_by_type on the same regression rows plus our own validation rows (choice, noul and score answers)" | |
| } | |