Jnx03's picture
mnemo2c-best180-v9target-step80: sft-lora
5ec59c2 verified
|
Raw
History Blame Contribute Delete
1.3 kB
---
library_name: peft
base_model: mistralai/Mistral-Nemo-Instruct-2407
license: apache-2.0
tags:
- thai
- kanitakorn
- single-model
- greedy-eval
language:
- th
- en
---
# Jnx03/kanitakorn-260614-mnemo2c-mnemo2c-best180-v9target-step80
Kanitakorn 2026-06-13 campaign artifact.
- Base: `mistralai/Mistral-Nemo-Instruct-2407`
- Version: `mnemo2c-best180-v9target-step80`
- Method: `sft-lora`
- Final-claim policy: one model, no BoN, no self-consistency, no ensemble, no model routing.
- Date: 2026-06-13
## Scores
| Benchmark | Score | Target | Status |
|---|---:|---:|---|
| aime24_th | pending | >15.0 | pending |
| aime24 | pending | >25.0 | pending |
| math500_th | pending | >56.0 | pending |
| math500 | pending | >82.0 | pending |
| livecodebench_th | pending | >35.0 | pending |
| livecodebench | pending | >60.0 | pending |
| openthaieval | pending | >80.0 | pending |
| hotpotqa_th_en | pending | >46.0 | pending |
| instruction_following_th_en | pending | >57.0 | pending |
| mt_bench_th_en | pending | >85.0 | pending |
| thaiexam | pending | >70.0 | pending |
| ifeval_th | pending | >82.0 | pending |
## Notes
Automatic checkpoint publish from runs/nonqwen-mistralnemo-v2c-continue-best180-v9target-r32-lr8e7-120/checkpoint-80. Single-model artifact; no BoN/self-consistency.