FastFlowLM commited on
Commit
070b3b9
·
verified ·
1 Parent(s): 62a857e

Create README.md

Browse files
Files changed (1) hide show
  1. README.md +69 -0
README.md ADDED
@@ -0,0 +1,69 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: mit
3
+ language:
4
+ - en
5
+ library_name: transformers
6
+ tags:
7
+ - qwen
8
+ - qwen3
9
+ - qwen3-8b
10
+ - text-generation
11
+ - AMD
12
+ - Ryzen
13
+ - NPU
14
+ pipeline_tag: text-generation
15
+ base_model:
16
+ - Qwen/Qwen3-7B-Instruct
17
+ ---
18
+
19
+ # 🐉 Qwen3 8B – Optimized for FastFlowLM on AMD Ryzen™ AI NPU (XDNA2 Only)
20
+
21
+ ## Model Summary
22
+ This model is based on **Qwen3 8B Instruct** from Alibaba Cloud. It uses the original Qwen3-7B architecture and weights, potentially with enhancements such as quantization or kernel-level acceleration for NPU efficiency via the FastFlowLM runtime.
23
+
24
+ > ✅ **Licensed under the permissive MIT License.**
25
+
26
+ ## 📝 License & Usage Terms
27
+
28
+ ### Base Model License
29
+ - Released under the MIT License by Alibaba Cloud:
30
+ 👉 https://huggingface.co/Qwen/Qwen3-7B-Instruct
31
+
32
+ - Permissions include:
33
+ - Free commercial and non-commercial use
34
+ - Permission to modify and redistribute
35
+ - Attribution not required, but appreciated
36
+
37
+ ### Redistribution Notice
38
+ - This repository **does not** contain original base weights.
39
+ - You must acquire the official weights from Qwen’s Hugging Face page:
40
+ 👉 https://huggingface.co/Qwen/Qwen3-7B-Instruct
41
+
42
+ ### If Fine-tuned
43
+ If this model has been modified (e.g., quantized, fine-tuned):
44
+
45
+ - **Base Model License**: MIT
46
+ - **Derivative Weights License**: [e.g., MIT, CC-BY-NC-4.0, custom]
47
+ - **Training Dataset License(s)**:
48
+ - [Dataset A] – [license]
49
+ - [Dataset B] – [license]
50
+
51
+ Ensure all datasets used are appropriately licensed for redistribution and use.
52
+
53
+ ## Intended Use
54
+ - **Recommended For**: High-performance on-device inference, private LLM workloads, local chat assistants, research
55
+ - **Not Recommended For**: Critical systems or regulated applications without thorough evaluation
56
+
57
+ ## Limitations & Risks
58
+ - Large model may require tuning for low-latency use
59
+ - Possibility of hallucination, toxicity, or factual errors
60
+ - May encode pretraining biases
61
+
62
+ ## Citation
63
+ ```bibtex
64
+ @misc{qwen32024,
65
+ title={Qwen3: Smaller, Smarter, and More Open},
66
+ author={Alibaba Cloud},
67
+ year={2024},
68
+ url={https://huggingface.co/Qwen}
69
+ ```