tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43

amr-fma training run.

  • Method: lora_sft
  • Base model: meta-llama/Llama-3.1-8B-Instruct
  • Dataset: DigitalLearningGmbH/MATH-lighteval (slug: math)
  • Seed: 43
  • Git commit: 8b979a30de6dfbf3b5a1052e42d8c0453b214d3f
  • Exp name: p1_sft_math_tooluse
  • WandB run: co3n04yu

Tags

  • phase:P1
  • domain:math

Checkpoints (branches)

  • step 1 β†’ revision step-00001
  • step 3 β†’ revision step-00003
  • step 6 β†’ revision step-00006
  • step 12 β†’ revision step-00012
  • step 23 β†’ revision step-00023
  • step 44 β†’ revision step-00044
  • step 83 β†’ revision step-00083
  • step 84 β†’ revision step-00084

Pin a specific checkpoint with revision=... in AutoModelForCausalLM.from_pretrained / PeftModel.from_pretrained.

Hyperparameter sections

checkpointing, dataset, evaluation, final_adapter_path, lora, model, optimization, prompt_style, runtime, sdpo, sequence, total_steps

Downloads last month
-
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43

Adapter
(2864)
this model