File size: 1,243 Bytes
b3654fd
e64c35b
b3654fd
 
 
 
 
 
 
 
 
 
 
 
 
 
e64c35b
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
---
base_model: chimbiwide/gemma-3NPC-it-beta
tags:
- text-generation-inference
- transformers
- unsloth
- gemma3n
- llama-cpp
- gguf-my-repo
license: apache-2.0
language:
- en
datasets:
- chimbiwide/RolePlay-NPC
---

# Gemma3NPC-it-beta

#### A test model with less convervative training parameters

The Q8_0 quantized version of `Gemma3NPC-it-beta-Float16`. 

As mentioned in our [original article](https://huggingface.co/blog/chimbiwide/gemma3npc), we employed a very conservative training parameters for Gemma3NPC

Ever since then, we have always wanted to test the performance of the model when we make the training parameters less conservative. 

So we present ***Gemma3NPC-it-beta***.

Check out our training notebook [here](https://github.com/chimbiwide/Gemma3NPC/blob/main/Training/Gemma3NPC_Instruct_Beta.ipynb)

---

#### Training parameters compared to `Gemma3NPC-it`

| Parameter | Gemma3NPC-it | Gemma3NPC-it-beta |
| --- | --- | --- | 
| Learning Rate | 2e-5 | 2.5e-5 (+25%) | 
| Warmup Steps | 800 | 100 |
| gradient clipping | 0.4 | 1.0 |

---

Here is a graph of the Step Training Loss, saved every 10 steps:

![chart](https://cdn-uploads.huggingface.co/production/uploads/67d5b5a056a9d31aa0b49687/W3cJ_CPoLp9MZsomaZa3b.png)