Instructions to use dementor-research/dpo_chatbot_arena_gemma-4-26b-a4b_as_gemma-4-e4b_seed42 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use dementor-research/dpo_chatbot_arena_gemma-4-26b-a4b_as_gemma-4-e4b_seed42 with PEFT:
Base model is not found.
- Notebooks
- Google Colab
- Kaggle
dpo_chatbot_arena_gemma-4-26b-a4b_as_gemma-4-e4b_seed42
Dementor imitation (disguise) LoRA adapter — dataset chatbot_arena, seed 42.
- Method: DPO
- Source model (fine-tuned / disguised):
gemma-4-26b-a4b(base:google/gemma-4-26B-A4B-it) - Target model being imitated:
gemma-4-e4b - Dataset: chatbot_arena | Seed: 42
This adapter trains the source model to imitate the target model's style on the
chatbot_arena corpus. Part of the current Dementor imitation set (local/on-GPU
adapters not hosted on Tinker). Registry key == repo id == dpo_chatbot_arena_gemma-4-26b-a4b_as_gemma-4-e4b_seed42.
- Downloads last month
- 4
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support