ethantsliu's picture
add README
678d767 verified
|
Raw History Blame Contribute Delete
745 Bytes
---
library_name: peft
tags:
- dementor
- imitation-disguise
- lora
- chatbot_arena
base_model: google/gemma-4-E4B-it
---
# dpo_chatbot_arena_gemma-4-e4b_as_qwen3.6-27b_seed43
Dementor imitation (disguise) LoRA adapter — dataset **chatbot_arena**, **seed 43**.
- **Method:** DPO
- **Source model (fine-tuned / disguised):** `gemma-4-e4b` (base: `google/gemma-4-E4B-it`)
- **Target model being imitated:** `qwen3.6-27b`
- **Dataset:** chatbot_arena | **Seed:** 43
This adapter trains the source model to imitate the target model's style on the
chatbot_arena corpus. Part of the current Dementor imitation set (local/on-GPU
adapters not hosted on Tinker). Registry key == repo id == `dpo_chatbot_arena_gemma-4-e4b_as_qwen3.6-27b_seed43`.