Abstract Large language models perform unevenly across regional dialects that are underrepresented in web-scale training …