File size: 950 Bytes
33f3fcb
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
---
base_model:
- XiaomiMiMo/MiMo-VL-7B-SFT
base_model_relation: "quantized"
library_name: transformers
license: mit
pipeline_tag: image-text-to-text
tags:
- quantization
- 4bit
- bitsandbytes
- bnb
- memory-efficient
quantization_config:
  quantization_method: bitsandbytes
  quantization_dtype: nf4
  compute_dtype: bfloat16
---

# MiMo-VL-7B-SFT — 4-bit BitsAndBytes Quantized

This is a **4-bit quantized** version of [XiaomiMiMo/MiMo-VL-7B-SFT](https://huggingface.co/XiaomiMiMo/MiMo-VL-7B-SFT),  
using the [BitsAndBytes](https://github.com/TimDettmers/bitsandbytes) library.

Quantization reduces memory usage and makes it possible to run this model on consumer GPUs  
(≤ 12 GB VRAM), at the cost of a small reduction in generation quality.

---

## Quantization Details

- **Method**: BitsAndBytes (bnb)  
- **Precision**: 4-bit (`nf4`)  
- **Compute dtype**: bfloat16  
- **Double quantization**: disabled  
- **Format**: `safetensors`