Phora68 commited on
Commit
f2571c4
·
verified ·
1 Parent(s): 1bcc8fd

Model card with training details and dataset provenance

Browse files
Files changed (1) hide show
  1. README.md +5 -22
README.md CHANGED
@@ -25,7 +25,7 @@ model-index:
25
 
26
  **Dr. Sage** is a fine-tuned version of [Qwen2.5-3B-Instruct](https://huggingface.co/Qwen/Qwen2.5-3B-Instruct) trained to act as a probing, honest, and empathetic therapeutic companion.
27
 
28
- Trained: **2026-03-19** | Base: `Qwen2.5-3B-Instruct` | Loss: `0.5166`
29
 
30
  ---
31
 
@@ -50,30 +50,13 @@ Dr. Sage never lectures. Never gives long speeches. Never asks more than **one q
50
 
51
  | Category | Samples |
52
  |----------|---------|
53
- | relationships | 675 |
54
- | general_therapeutic | 596 |
55
- | depression | 558 |
56
- | anxiety | 544 |
57
- | family_dynamics | 414 |
58
- | trauma | 333 |
59
- | work_burnout | 251 |
60
- | anger | 250 |
61
- | loneliness | 243 |
62
- | substance_avoidance | 242 |
63
- | shame_esteem | 232 |
64
- | grief_loss | 230 |
65
- | identity_transition | 229 |
66
- | self_esteem | 225 |
67
- | boundaries_pleasing | 225 |
68
- | crisis_safety | 30 |
69
- | physical_somatic | 10 |
70
 
71
  ### By source
72
 
73
  | Source | Samples |
74
  |--------|---------|
75
- | synthetic | 4,375 |
76
- | alpaca | 912 |
77
 
78
  Full dataset available at [Phora68/dr-sage-dataset](https://huggingface.co/datasets/Phora68/dr-sage-dataset)
79
 
@@ -95,8 +78,8 @@ Full dataset available at [Phora68/dr-sage-dataset](https://huggingface.co/datas
95
  | LR schedule | cosine |
96
  | Optimizer | adamw_8bit |
97
  | Hardware | A100 80GB |
98
- | Final loss | 0.5166 |
99
- | Training time | 15 min |
100
 
101
  ---
102
 
 
25
 
26
  **Dr. Sage** is a fine-tuned version of [Qwen2.5-3B-Instruct](https://huggingface.co/Qwen/Qwen2.5-3B-Instruct) trained to act as a probing, honest, and empathetic therapeutic companion.
27
 
28
+ Trained: **2026-03-19** | Base: `Qwen2.5-3B-Instruct` | Loss: `0.0000`
29
 
30
  ---
31
 
 
50
 
51
  | Category | Samples |
52
  |----------|---------|
53
+ | general_therapeutic | 5,287 |
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
54
 
55
  ### By source
56
 
57
  | Source | Samples |
58
  |--------|---------|
59
+ | synthetic + alpaca | 5,287 |
 
60
 
61
  Full dataset available at [Phora68/dr-sage-dataset](https://huggingface.co/datasets/Phora68/dr-sage-dataset)
62
 
 
78
  | LR schedule | cosine |
79
  | Optimizer | adamw_8bit |
80
  | Hardware | A100 80GB |
81
+ | Final loss | 0.0000 |
82
+ | Training time | 0 min |
83
 
84
  ---
85