Instructions to use t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16 with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir Qwen3.5-2B-MLX-MTP-bf16 t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
Download mtplx_runtime.json from t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16: direct link, hf CLI and curl.
- Browser
- Download file 392 Bytes
-
https://huggingface.co/t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16/resolve/eb14195eea53195dc00d934d1eb691cde5e243e1/mtplx_runtime.json
- Command line
-
hf download hf://t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16@eb14195eea53195dc00d934d1eb691cde5e243e1/mtplx_runtime.json
-
curl -L -o mtplx_runtime.json https://huggingface.co/t0rr3sp3dr0/Qwen3.5-2B-MLX-MTP-bf16/resolve/eb14195eea53195dc00d934d1eb691cde5e243e1/mtplx_runtime.json
392 Bytes
| { | |
| "arch_id": "qwen3-next-mtp", | |
| "mlx_lm_extra_tensors": { | |
| "mtp_file": "mtp/weights.safetensors" | |
| }, | |
| "mtp_contract": { | |
| "base_hidden_variant": "post_norm", | |
| "concat_order": "embedding_hidden", | |
| "hidden_variant": "post_norm", | |
| "mtp_position_mode": "local", | |
| "mtp_quant_group_size": 64, | |
| "mtp_quant_mode": "affine" | |
| }, | |
| "mtp_depth_max": 3, | |
| "mtp_sidecar": "bf16" | |
| } |