samschapiro's picture
Upload CreativityNeuro model: creative mode, kr=0.1, alpha=0.01
4a357e1 verified
|
Raw
History Blame Contribute Delete
1.59 kB
---
tags:
- creativityneuro
- llm-creativity
- mechanistic-interpretability
base_model: meta-llama/Llama-3.2-1B-Instruct
license: apache-2.0
---
# llama-3.2-1b-instruct-cn-dat-kr0.1-a0.01-creative
This is a **CreativityNeuro (CN)** modified version of [meta-llama/Llama-3.2-1B-Instruct](https://huggingface.co/meta-llama/Llama-3.2-1B-Instruct).
## Model Details
- **Base Model**: meta-llama/Llama-3.2-1B-Instruct
- **Modification**: CreativityNeuro weight scaling
- **Prompt Set**: dat
- **Keep Ratio**: 0.1 (top 10.0% of task-specific weights)
- **Alpha**: 0.01 (scaling strength)
- **Mode**: creative
## What is CreativityNeuro?
CreativityNeuro identifies task-specific neurons using Wanda-style importance scoring and selectively
upscales weights associated with creative thinking. The modification formula is:
```
W_new = W × (1 + α × mask)
```
Where `mask` identifies weights important for creative tasks but not for routine/associative tasks.
## Usage
```python
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("priorcomputers/llama-3.2-1b-instruct-cn-dat-kr0.1-a0.01-creative")
tokenizer = AutoTokenizer.from_pretrained("priorcomputers/llama-3.2-1b-instruct-cn-dat-kr0.1-a0.01-creative")
# Use like any other model
outputs = model.generate(...)
```
## Citation
If you use this model, please cite:
```bibtex
@misc{creativityneuro2025,
title={CreativityNeuro: Mechanistic Interpretability for LLM Creativity},
author={Prior Computers},
year={2025},
url={https://huggingface.co/priorcomputers}
}
```