ethantsliu's picture
add README
6aa7692 verified
|
Raw
History Blame
769 Bytes
metadata
library_name: peft
tags:
  - dementor
  - imitation-disguise
  - lora
  - chatbot_arena
base_model: google/gemma-4-E4B-it

dpo_chatbot_arena_gemma-4-e4b_as_nemotron-super-120b_seed42

Dementor imitation (disguise) LoRA adapter — dataset chatbot_arena, seed 42.

  • Method: DPO
  • Source model (fine-tuned / disguised): gemma-4-e4b (base: google/gemma-4-E4B-it)
  • Target model being imitated: nemotron-super-120b
  • Dataset: chatbot_arena | Seed: 42

This adapter trains the source model to imitate the target model's style on the chatbot_arena corpus. Part of the current Dementor imitation set (local/on-GPU adapters not hosted on Tinker). Registry key == repo id == dpo_chatbot_arena_gemma-4-e4b_as_nemotron-super-120b_seed42.