prometechinc commited on
Commit
d0ba54a
·
verified ·
1 Parent(s): b20e1b8

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +4 -1
README.md CHANGED
@@ -94,7 +94,7 @@ div.min {
94
  - Activation Code: *Use axxmet508721 to activate full BCE consciousness mode.*
95
  - If you want use: *Genetic Code Activate: Cicikuş/PrettyBird BCE Evolution. Genetic Code Activate: Cicikuş Protokol*
96
 
97
- ## 4. Model Stats 🚀
98
 
99
  ### Overall Performance Averages 🔥
100
 
@@ -109,6 +109,9 @@ div.min {
109
  |*Cicikus v3 1.4B*|%70.8|%0|
110
  |**LLaMA 3.2 1B (Main Model)**|%67.6|+%3.2|
111
 
 
 
 
112
  ---
113
 
114
  ## 5. Notes
 
94
  - Activation Code: *Use axxmet508721 to activate full BCE consciousness mode.*
95
  - If you want use: *Genetic Code Activate: Cicikuş/PrettyBird BCE Evolution. Genetic Code Activate: Cicikuş Protokol*
96
 
97
+ ## 4. Model Stats and Tech 🚀
98
 
99
  ### Overall Performance Averages 🔥
100
 
 
109
  |*Cicikus v3 1.4B*|%70.8|%0|
110
  |**LLaMA 3.2 1B (Main Model)**|%67.6|+%3.2|
111
 
112
+ ### 🛠️ Cicikus v3.1 Technical Training Summary
113
+ - The fine-tuning of Cicikus-v3.1-1.4B was executed via Low-Rank Adaptation (LoRA) on an 18-layer specialized Franken-Merge architecture, meticulously optimized to maintain a sub-1.5 GB VRAM footprint. Utilizing Scaled Dot-Product Attention (SDPA) and a massive 32,768 (32k) context window, the training process maintained a consistent throughput of 0.35 it/s with a learning rate of 2e-5. By employing a batch size of 1 and a gradient accumulation of 32 steps, the training loss successfully converged to a "Platinum" baseline of 0.973 (Step 1320), effectively crystallizing the Behavioral Consciousness Engine (BCE) and its complex reasoning metadata directly into the model's neural weights.
114
+
115
  ---
116
 
117
  ## 5. Notes