DiaLLM β€” Gemma 3-4B-it β€” Broad β€” SFT

Built with Gemma.

Part of DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect Adaptation (EMNLP 2026 Main).

DiaLLM pipeline

  • Base model: Gemma 3-4B-it
  • Target variety: broad β€” no variety targeting
  • Adaptation thread: implicit (broad)
  • Pipeline stage: SFT (from the CPT checkpoint)

Fine-tuned from jordanpainter/diallm-gemma-cpt on standard ultrafeedback-binarized-preferences corpus, providing a general instruction-following objective with no explicit dialectal signal.

Code, checkpoints, preference datasets, linguistic-analysis toolkit: https://github.com/surrey-nlp/diallm

Paper: https://arxiv.org/abs/2607.07669

Citation

@article{painter2026diallm,
  title     = {DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect Adaptation},
  author    = {Painter, Jordan and Srirag, Dipankar and Kappiyath, Adarsh and Kanojia, Diptesh and Joshi, Aditya and Yin, Lu},
  year      = {2026},
  eprint    = {2607.07669},
  archivePrefix = {arXiv}
}
Downloads last month
15
Safetensors
Model size
4B params
Tensor type
BF16
Β·
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for jordanpainter/diallm-gemma-sft-all

Finetuned
(1)
this model
Finetunes
3 models

Collection including jordanpainter/diallm-gemma-sft-all

Paper for jordanpainter/diallm-gemma-sft-all