Papers
arxiv:2607.07669

DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect Adaptation

Published on Jul 8
Authors:
,
,
,
,
,

Abstract

DiaLLM compares continual pretraining and alignment strategies for generating Australian, Indian, and Northern British English, finding that dialectal generation and benchmark robustness are dissociated and that explicit variety-targeted adaptation improves dialectal recognition but reveals a gap between reward optimization and human preference.

Large language models increasingly understand dialectal English, yet still produce only standard, US-leaning English, leaving dialectal generation, the harder half of the problem, largely unaddressed. We introduce DiaLLM, which continually pretrains three open-weight language model families on the International Corpus of English and applies implicit and explicit post-training paradigms, each combined with three model alignment strategies, giving the first controlled comparison of these components across Australian, Indian, and Northern British English. Our results reveal that dialectal robustness and generation are dissociated: benchmarks are shaped by continual pretraining and SFT, while alignment visibly reshapes generation in ways benchmarks do not capture. Explicit variety-targeted adaptation produces output reliably recognised as dialectal and preferred over broad alignment, yet the method that most aggressively optimises the dialectal reward is not preferred by human evaluators. Independent linguistic analysis corroborates this reward-quality gap, most clearly on two of the three families. No single alignment method dominates, and closing the gap will require richer reward designs and continued investment in dialectal resources. We release all code, checkpoints, and preference datasets.

Community

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2607.07669
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 23

Browse 23 models citing this paper

Datasets citing this paper 4

Spaces citing this paper 0

No Space linking this paper

Cite arxiv.org/abs/2607.07669 in a Space README.md to link it from this page.

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.