nick-sh-oh's picture
Link paper: arXiv:2608.05224
0e8aa55 verified
|
Raw History Blame Contribute Delete
5.61 kB
metadata
license: llama3.1
datasets:
  - marcelbinz/Psych-101
language:
  - en
base_model:
  - unsloth/Llama-3.1-8B
base_model_relation: adapter
pipeline_tag: text-generation
library_name: peft
tags:
  - psychology
  - cognitive science
  - human behavior
  - unsloth
  - lora
Llama-Centaur-8B-LoRA-r4

Meta Llama

socius Paper Parameters LoRA Dataset

Llama-Centaur-8B-LoRA-r4

LoRA adapter for Llama-Centaur-8B, fine-tuned on the full Psych-101 as part of the LoRA-rank sweep and dataset-size ablation for Small Foundation Models of Human Cognition and Behaviour.

field value
base model unsloth/Llama-3.1-8B
LoRA rank 4 (alpha = rank, rsLoRA)
data fraction 100% of Psych-101
training 1 epoch, completion-only loss, seed 3407

Load with PEFT on top of unsloth/Llama-3.1-8B, or evaluate with the project's eval_model.py --backend unsloth.