FLEURS-Trigram Whisper Base No-Language Checkpoint

Summary

This repository contains a Whisper checkpoint for Chichewa/Nyanja automatic speech recognition, fine-tuned from openai/whisper-base.

  • Experiment type: fleurs-trigram
  • Base model: openai/whisper-base
  • Training condition: no_language
  • Release artifact: full fine-tuned checkpoint selected from the best training checkpoint

Intended use

This checkpoint is intended for research and evaluation on Chichewa/Nyanja ASR. It is not a production-ready speech system and should be validated carefully before downstream use.

How to use

This repository contains a full fine-tuned checkpoint. It can be loaded directly with Transformers.

from transformers import AutoModelForSpeechSeq2Seq, AutoProcessor

model_id = "ai4good-labyrinth/fleurs-trigram-hours1p00-whisper-base-no-language"
processor = AutoProcessor.from_pretrained(model_id)
model = AutoModelForSpeechSeq2Seq.from_pretrained(model_id)

Training data

  • Training source: FLEURS train + Chichewa Trigrams train
  • Evaluation source during training: FLEURS dev
  • Train examples before duration filtering: 28629
  • Train examples after duration filtering: 28588
  • Dev examples before duration filtering: 311
  • Dev examples after duration filtering: 305
  • Duration filter used during training: min_duration_seconds=0.0, max_duration_seconds=30.0

Training procedure

  • Fine-tuning script: experiments/whisper_finetune/finetune_whisper.py
  • Base model: openai/whisper-base
  • Task: transcribe
  • Language hint during training/evaluation: none, corresponding to --language auto in standalone evaluation
  • Mixed precision: no
  • Gradient checkpointing: False
  • Selected checkpoint step: 5000
  • Selected checkpoint epoch: 1.40

Training-time dev selection

The best checkpoint was selected using trainer-side dev evaluation on the duration-filtered FLEURS dev split.

  • Dev WER: 0.6158
  • Dev CER: 0.2003
  • Dev loss: 0.8416

These values come from the training pipeline and may differ slightly from standalone post-hoc evaluation because the decoding path is not perfectly identical.

Evaluation protocol

Standalone evaluation is recommended for the final release. Filtered and unfiltered results should be reported separately.

  • Filtered evaluation: min_duration_seconds=0, max_duration_seconds=30
  • Unfiltered evaluation: no duration constraint
  • Decoding task: transcribe
  • Language hint: auto

Evaluation summary

Dataset Split Setting Num examples WER CER Notes
FLEURS dev filtered 305 0.6248 0.2223 Filtered to 30 seconds
FLEURS dev unfiltered TBD TBD TBD Standalone eval pending
FLEURS test filtered 745 0.6973 0.2457 Filtered to 30 seconds
FLEURS test unfiltered TBD TBD TBD Standalone eval pending
Zambezi dev filtered 613 0.8684 0.3226 Filtered to 30 seconds
Zambezi dev unfiltered TBD TBD TBD Standalone eval pending
Zambezi test filtered 427 6173.0000 42187.0000 Filtered to 30 seconds
Zambezi test unfiltered TBD TBD TBD Standalone eval pending

Files in this repository

  • Model weights and config: repository root
  • Processor/tokenizer files: repository root
  • Evaluation JSON files: eval/...

Known limitations

  • Whisper does not provide an official Nyanja/Chichewa language token.
  • Users must also comply with the upstream dataset licenses and any upstream model license obligations.
  • Standalone evaluation and trainer-side evaluation can differ slightly even on the same split and duration filter.
  • Cross-dataset results should be interpreted carefully because transcription conventions may differ across corpora.

Citation

If you use this checkpoint, please cite:

  • the Whisper paper
  • the FLEURS dataset
  • this repository
@misc{fleurs_trigram_hours1p00_whisper_base_no_language_2026,
  title        = {FLEURS-Trigram Whisper Base No-Language Checkpoint},
  author       = {AI4Good Labyrinth Team},
  year         = {2026},
  howpublished = {\url{https://huggingface.co/ai4good-labyrinth/fleurs-trigram-hours1p00-whisper-base-no-language}},
  note         = {Whisper fine-tuning for Chichewa/Nyanja ASR}
}
Downloads last month
15
Safetensors
Model size
72.6M params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for ai4good-labyrinth/fleurs-trigram-hours1p00-whisper-base-no-language

Finetuned
(762)
this model