ggerganov HF Staff commited on
Commit
49ecdb2
·
verified ·
1 Parent(s): cac6e14

Upload folder using huggingface_hub

Browse files
.gitattributes CHANGED
@@ -42,3 +42,4 @@ mmproj-MiMo-V2.6-Flash-RL-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
42
  mtp-MiMo-V2.6-Flash-RL-BF16.gguf filter=lfs diff=lfs merge=lfs -text
43
  mtp-MiMo-V2.6-Flash-RL-MXFP4.gguf filter=lfs diff=lfs merge=lfs -text
44
  mtp-MiMo-V2.6-Flash-RL-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
 
 
42
  mtp-MiMo-V2.6-Flash-RL-BF16.gguf filter=lfs diff=lfs merge=lfs -text
43
  mtp-MiMo-V2.6-Flash-RL-MXFP4.gguf filter=lfs diff=lfs merge=lfs -text
44
  mtp-MiMo-V2.6-Flash-RL-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
45
+ mtp-MiMo-V2.6-Flash-RL-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
MiMo-V2.6-Flash-RL-Q2_K-00002-of-00002.gguf CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:039c408dafc5b9507fc8f4c35a3939610825b33647cb1446f2dd6a39419dd58f
3
- size 125711620192
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:218634f8e626c06998350a0b23f0b64fc467a2750e84ca9a745c0a3eae2ec7f0
3
+ size 126101690464
README.md CHANGED
@@ -22,7 +22,7 @@ llama serve -hf ggml-org/MiMo-V2.6-Flash-RL-GGUF
22
  ### Notes
23
  - The MXFP4 output keeps the routed experts at their native MXFP4 precision.
24
  - The Q2_K output keeps the expert down projections at MXFP4, and quantizes the gate/up projections to Q2_K.
25
- - Includes MTP sidecars (MXFP4 and Q8_0) for speculative decoding (`--mtp`).
26
  - Includes a Q8_0 mmproj for the vision and audio encoders.
27
  - Currently, the Q2 models do not use an imatrix calibration due to lack of one.
28
 
 
22
  ### Notes
23
  - The MXFP4 output keeps the routed experts at their native MXFP4 precision.
24
  - The Q2_K output keeps the expert down projections at MXFP4, and quantizes the gate/up projections to Q2_K.
25
+ - Includes MTP sidecars (Q4_0 and Q8_0) for speculative decoding (`--mtp`).
26
  - Includes a Q8_0 mmproj for the vision and audio encoders.
27
  - Currently, the Q2 models do not use an imatrix calibration due to lack of one.
28
 
convert.log CHANGED
The diff for this file is too large to render. See raw diff
 
mtp-MiMo-V2.6-Flash-RL-Q4_0.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:e7dc7917bd8f824179d827fb7f47fa3aac4df8a439bcd24a01df2761df424787
3
+ size 1264898240