File size: 1,429 Bytes
0274562
 
4c18248
 
 
 
 
 
 
 
 
0274562
 
a1f901d
 
942e2bd
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
9c9a3fc
 
 
8745586
 
 
 
 
 
 
9c9a3fc
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
---
library_name: transformers
license: apache-2.0
datasets:
- kurakurai/luth-sft
language:
- fr
- en
base_model:
- Qwen/Qwen3-0.6B
pipeline_tag: text-generation
---

![Kurakura AI Logo](media/logo_kurakura.png)

# Luth-0.6B

**Luth-0.6B** is a French fine-tuned version of [Qwen3-0.6B](https://huggingface.co/Qwen/Qwen3-0.6B), trained on the [Luth-SFT](https://huggingface.co/datasets/kurakurai/luth-sft) dataset. The model has drastically improved its French capabilities in instruction following, math, and general knowledge. Additionally, its English capabilities have remained stable and have even increased in some areas.

## Model Details

Luth-0.6B was trained using full fine-tuning on the Luth-SFT dataset with [Axolotl](https://github.com/axolotl-ai-cloud/axolotl). The resulting model was then merged with the base Qwen3-0.6B model. This process successfully retained the model's English capabilities while improving its performance on nearly all benchmarks in both French and English.

## Benchmark Results

**French Evaluation:**

![French Evaluation](media/french_evaluation.png)

**English Evaluation:**

![English Evaluation](media/english_evaluation.png)

## Citation

```bibtex
@misc{luth2025kurakurai,
  title   = {Luth-0.6B},
  author  = {Kurakura AI Team},
  year    = {2025},
  howpublished = {\url{https://huggingface.co/kurakurai/Luth-0.6B}},
  note    = {Qwen3-0.6B fine-tuned on French datasets}
}
```