bullerwins commited on
Commit
6c09f6a
·
verified ·
1 Parent(s): ef5f89e

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +12 -6
README.md CHANGED
@@ -5,13 +5,19 @@ license_name: minimax-community
5
  license_link: LICENSE
6
  library_name: transformers
7
  tags:
8
- - multimodal
9
- - moe
10
- - agent
11
- - coding
12
- - video
 
 
13
  ---
14
 
 
 
 
 
15
  <div align="center">
16
  <img width="60%" src="figures/logo.svg" alt="MiniMax">
17
  </div>
@@ -84,4 +90,4 @@ We recommend the following parameters for best performance: `temperature=1.0`, `
84
 
85
  ## Contact Us
86
 
87
- Contact us at [model@minimax.io](mailto:model@minimax.io).
 
5
  license_link: LICENSE
6
  library_name: transformers
7
  tags:
8
+ - multimodal
9
+ - moe
10
+ - agent
11
+ - coding
12
+ - video
13
+ base_model:
14
+ - MiniMaxAI/MiniMax-M3
15
  ---
16
 
17
+ Experimental int4 w4a16, I have not been able to test it as the vLLM M3's support PR does not support pipeline paralelism and I don't have the hardware to test tensor paralelism, so here may be dragons, but people like to tinker.
18
+ You will need this PR from vllm to make it work https://github.com/vllm-project/vllm/pull/45381
19
+ This is using RTN quantization, not full calibrated Auto-round.
20
+
21
  <div align="center">
22
  <img width="60%" src="figures/logo.svg" alt="MiniMax">
23
  </div>
 
90
 
91
  ## Contact Us
92
 
93
+ Contact us at [model@minimax.io](mailto:model@minimax.io).