--- library_name: peft base_model: meta-llama/Llama-3.1-8B-Instruct tags: - amr-fma - lora_sft - domain:math - phase:P1 --- # tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43 amr-fma training run. - **Method**: `lora_sft` - **Base model**: `meta-llama/Llama-3.1-8B-Instruct` - **Dataset**: `DigitalLearningGmbH/MATH-lighteval` (slug: `math`) - **Seed**: `43` - **Git commit**: `8b979a30de6dfbf3b5a1052e42d8c0453b214d3f` - **Exp name**: `p1_sft_math_tooluse` - **WandB run**: `co3n04yu` ## Tags - phase:P1 - domain:math ## Checkpoints (branches) - step 1 → revision `step-00001` - step 3 → revision `step-00003` - step 6 → revision `step-00006` - step 12 → revision `step-00012` - step 23 → revision `step-00023` - step 44 → revision `step-00044` - step 83 → revision `step-00083` - step 84 → revision `step-00084` Pin a specific checkpoint with `revision=...` in `AutoModelForCausalLM.from_pretrained` / `PeftModel.from_pretrained`. ## Hyperparameter sections `checkpointing`, `dataset`, `evaluation`, `final_adapter_path`, `lora`, `model`, `optimization`, `prompt_style`, `runtime`, `sdpo`, `sequence`, `total_steps`