Jnx03's picture
mnemo2c-best180-v9target-step80: sft-lora
5ec59c2 verified
|
Raw
History Blame Contribute Delete
1.3 kB
metadata
library_name: peft
base_model: mistralai/Mistral-Nemo-Instruct-2407
license: apache-2.0
tags:
  - thai
  - kanitakorn
  - single-model
  - greedy-eval
language:
  - th
  - en

Jnx03/kanitakorn-260614-mnemo2c-mnemo2c-best180-v9target-step80

Kanitakorn 2026-06-13 campaign artifact.

  • Base: mistralai/Mistral-Nemo-Instruct-2407
  • Version: mnemo2c-best180-v9target-step80
  • Method: sft-lora
  • Final-claim policy: one model, no BoN, no self-consistency, no ensemble, no model routing.
  • Date: 2026-06-13

Scores

Benchmark Score Target Status
aime24_th pending >15.0 pending
aime24 pending >25.0 pending
math500_th pending >56.0 pending
math500 pending >82.0 pending
livecodebench_th pending >35.0 pending
livecodebench pending >60.0 pending
openthaieval pending >80.0 pending
hotpotqa_th_en pending >46.0 pending
instruction_following_th_en pending >57.0 pending
mt_bench_th_en pending >85.0 pending
thaiexam pending >70.0 pending
ifeval_th pending >82.0 pending

Notes

Automatic checkpoint publish from runs/nonqwen-mistralnemo-v2c-continue-best180-v9target-r32-lr8e7-120/checkpoint-80. Single-model artifact; no BoN/self-consistency.