Update README.md
Browse files
README.md
CHANGED
|
@@ -94,7 +94,7 @@ To deploy the WINT2 quantized version using FastDeploy on two 80G GPUs, run the
|
|
| 94 |
|
| 95 |
```bash
|
| 96 |
python -m fastdeploy.entrypoints.openai.api_server \
|
| 97 |
-
--model "baidu/ERNIE-4.5-300B-A47B-
|
| 98 |
--port 8180 \
|
| 99 |
--metrics-port 8181 \
|
| 100 |
--engine-worker-queue-port 8182 \
|
|
|
|
| 94 |
|
| 95 |
```bash
|
| 96 |
python -m fastdeploy.entrypoints.openai.api_server \
|
| 97 |
+
--model "baidu/ERNIE-4.5-300B-A47B-2Bits-TP2-Paddle" \
|
| 98 |
--port 8180 \
|
| 99 |
--metrics-port 8181 \
|
| 100 |
--engine-worker-queue-port 8182 \
|