File size: 633 Bytes
5a42749
 
 
 
06d8e28
 
5a42749
 
 
 
48762dc
053810f
cb0d7d2
322a63b
27371d1
322a63b
27371d1
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
---
license: apache-2.0
tags:
- awq
- AMD
- Ryzen AI
---

# Optmized LLM models for Ryzen AI.

The models are quantized & tested via [Ryzen AI SW 1.2](https://github.com/amd/RyzenAI-SW/tree/main/example/transformers).

The folder `quantized_models` contains a set of LLMs quantized via different algorithms.    
The folder `onnx_model_nodes` contains the `mlp.up_proj` nodes for `-npugpu` device parallel config. 

Due to the disordered Ryzen AI Python environment, you may need to install different versions of transformers:
```python
transformers==4.37.2  # Others
transformers==4.39.1  # GEMMA
transformers>=4.42  # Mistral
```