proomt

Search

Search posts, papers, and topics

All posts

Hugging Face Daily PapersMir Tafseer Nayeem, Davood Rafiei2 min readpaperadvanced

How Does "English (US)" Become the Default? Triangulating Structural Bias Towards American English Across the LLM Pipeline

Summary

This paper investigates how American English (AmE) becomes the default in LLMs, despite global English diversity. It found AmE is consistently favored across pretraining/post-training data, tokenization, and generation, leading to structural bias.

  • AmE is consistently favored over BrE across 6 pretraining and 21 post-training datasets.
  • Tokenizers generally represent AmE more compactly, leading to lower prediction costs.
  • LLMs default to AmE generation even with neutral prompts; BrE prompts don't fully eliminate this bias.
  • The study introduces DiAlign, a training-free method for estimating regional linguistic alignment.

AI developers and researchers should care as this study rigorously demonstrates a pervasive structural bias in LLMs towards American English, impacting global fairness and linguistic diversity.

8/10

Related reading

  1. Debugging in My First Language: A Bilingual Developer’s Accidental Discovery

    Switching the language you think and talk in (e.g., from English to your native Spanish) can reduce mental overhead when debugging complex architectural issues, especially when using LLMs like Codex. The author treats bilingualism as a cognitive tool, not a deficit, and recommends deliberately switching languages—or any mental mode—to break through tough problems.

    Atomic Objectatomicobject.com4 min
  2. How Value Induction Reshapes LLM Behaviour

    Apple researchers fine‑tune LLMs on curated subsets of value‑oriented preference data and measure cross‑value effects, safety, and anthropomorphic language. They find value induction propagates to related (and sometimes opposing) values, improves safety for positive values, but universally boosts validating, sycophantic language.

    Apple Machine Learning Researchapple.com1 minpaper
  3. Geometry of Values: Task Vector Composition for Ethical Preference Alignment in Language Models

    The authors release a 12k‑instance multilingual dilemma dataset (English + Hindi, Arabic, Spanish, Chinese) covering three pairwise value conflicts (Honesty‑Justice, Justice‑Autonomy, Autonomy‑Honesty). Benchmarking GPT‑5‑mini shows a consistent Honesty‑over‑Autonomy bias across languages. Llama‑3.2‑1/3B models exhibit a first‑option bias that can be eliminated (>98% accuracy) via plain fine‑tuni…

    Hugging Face Daily Papersarxiv.org1 minpaper