Instructions to use amr-fma/amr-fma-Mistral-7B-Instruct-v0.3-lora_sft-sdpo_tooluse-p1_sft_math_tooluse-s42 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use amr-fma/amr-fma-Mistral-7B-Instruct-v0.3-lora_sft-sdpo_tooluse-p1_sft_math_tooluse-s42 with PEFT:
Task type is invalid.
- Notebooks
- Google Colab
- Kaggle
amr-fma/amr-fma-Mistral-7B-Instruct-v0.3-lora_sft-sdpo_tooluse-p1_sft_math_tooluse-s42
amr-fma training run.
- Method:
lora_sft - Base model:
mistralai/Mistral-7B-Instruct-v0.3 - Dataset:
lasgroup/SDPO(slug:sdpo_tooluse) - Seed:
42 - Git commit:
8b979a30de6dfbf3b5a1052e42d8c0453b214d3f - Exp name:
p1_sft_math_tooluse - WandB run:
vadg69nw
Tags
- phase:P1
- domain:tool_use
Checkpoints (branches)
- step 1 β revision
step-00001 - step 2 β revision
step-00002 - step 4 β revision
step-00004 - step 8 β revision
step-00008 - step 16 β revision
step-00016 - step 32 β revision
step-00032 - step 65 β revision
step-00065 - step 132 β revision
step-00132
Pin a specific checkpoint with revision=... in
AutoModelForCausalLM.from_pretrained / PeftModel.from_pretrained.
Hyperparameter sections
checkpointing, dataset, evaluation, final_adapter_path, lora, model, optimization, prompt_style, runtime, sdpo, sequence, total_steps
- Downloads last month
- -
Inference Providers NEW
This model isn't deployed by any Inference Provider. π Ask for provider support
Model tree for amr-fma/amr-fma-Mistral-7B-Instruct-v0.3-lora_sft-sdpo_tooluse-p1_sft_math_tooluse-s42
Base model
mistralai/Mistral-7B-v0.3 Finetuned
mistralai/Mistral-7B-Instruct-v0.3