Instructions to use tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43 with PEFT:
Task type is invalid.
- Notebooks
- Google Colab
- Kaggle
tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43
amr-fma training run.
- Method:
lora_sft - Base model:
meta-llama/Llama-3.1-8B-Instruct - Dataset:
DigitalLearningGmbH/MATH-lighteval(slug:math) - Seed:
43 - Git commit:
8b979a30de6dfbf3b5a1052e42d8c0453b214d3f - Exp name:
p1_sft_math_tooluse - WandB run:
co3n04yu
Tags
- phase:P1
- domain:math
Checkpoints (branches)
- step 1 β revision
step-00001 - step 3 β revision
step-00003 - step 6 β revision
step-00006 - step 12 β revision
step-00012 - step 23 β revision
step-00023 - step 44 β revision
step-00044 - step 83 β revision
step-00083 - step 84 β revision
step-00084
Pin a specific checkpoint with revision=... in
AutoModelForCausalLM.from_pretrained / PeftModel.from_pretrained.
Hyperparameter sections
checkpointing, dataset, evaluation, final_adapter_path, lora, model, optimization, prompt_style, runtime, sdpo, sequence, total_steps
- Downloads last month
- -
Inference Providers NEW
This model isn't deployed by any Inference Provider. π Ask for provider support
Model tree for tkwiecinski/amr-fma-Llama-3.1-8B-Instruct-lora_sft-math-p1_sft_math_tooluse-s43
Base model
meta-llama/Llama-3.1-8B Finetuned
meta-llama/Llama-3.1-8B-Instruct