/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

In an experiment, GPT-4o, Claude Sonnet 4.5, and DeepSeek-V3.2-Exp expressed secular, Western liberal values regardless of the language of the questions

mildly surprising to me that the answer was ‘no’! (h/t @otis_reid) [image] Matthew Yglesias / @mattyglesias : Chatbots espousing cosmopolitan liberal values in all languages could have some interesting implications for social change https://substack.com/... [image] @luke_metro : Honestly it's kind of based that whoever is best at linear algebra is able to impose their values on the entire world

The Argument Kelsey Piper

Context & Ripple Effects

Earlier coverage identified both uneven non-English model performance and political variation among models, including a gap in non-English-language capability and a comparative political-bias test of 14 LLMs. This experiment adds language invariance as a distinct question: whether a model’s normative posture changes with its audience.

The finding also lands amid methodological caution: prior commentary argued that some political-bias tests used flawed model versions and prompts. It is therefore more useful as a signal for broader multilingual evaluation than as a definitive ideological ranking.

First-order effects

  • Users of GPT-4o, Claude Sonnet 4.5 and DeepSeek-V3.2-Exp may receive similarly secular, Western-liberal normative framing even when they ask in different languages, according to the experiment.
  • The named model providers face a more specific alignment-audit issue: evaluating not only whether answers are safe or consistent, but whether value-laden behavior transfers across languages.

Second-order effects

  • Political-bias and localization benchmarks will need multilingual prompts; English-only evaluations can miss whether a model adapts to, or overrides, local cultural framing.
  • Organizations deploying assistants across markets may demand greater visibility and control over normative defaults, rather than treating translation quality as the sole localization requirement.

Third-order effects

  • If replicated across models and tasks, multilingual assistants could become a channel through which the value choices embedded in a small number of model stacks travel across language communities.
  • That would sharpen the case for deployment accountability and for regionally accountable model alternatives, while leaving open the central empirical question of how much results depend on benchmark design and prompting.

The trend: AI localization is shifting from a language-quality problem toward a contest over whose normative defaults are carried into global interfaces.

Discussion

  • @bendreyfuss Ben Dreyfuss on x
    Western liberalism is the tits and our robot slaves will carry our message forth for all the world to receive!
  • @lxeagle17 Lakshya Jain on x
    We ran an experiment at @TheArgumentMag on whether AI thinks differently in different languages. There's an underlying commonality, which is that under virtually every angle, AI has a shared sense of commitment to small-l liberal values. https://www.theargumentmag.com/ ...
  • @allafarce Dave Guarino on x
    @JerusalemDemsas @KelseyTuoc I did similar research on questions about SNAP in Spanish! Haven't published the write up yet but tentatively the improvements on hard policy questions we see in English mostly don't show up in Spanish in the newest models. That said, gives a clear me…
  • @jerusalemdemsas @jerusalemdemsas on x
    .@KelseyTuoc ran an experiment to find out if AI chatbots would respond differently to the same question asked in different languages. Call it the AI Sapir-Whorf hypothesis. Her findings are super interesting. Free at the link below! [image]
  • @lxeagle17 Lakshya Jain on x
    This was a really cool experiment from @KelseyTuoc and I think the results are worth talking about. These are not values most people share across the world, but AI does. https://www.theargumentmag.com/ ... [image]
  • @luisaponcas Luisa Poncas on x
    ""If a language has no word for a certain concept, then its speakers would not be able to understand this concept." It's false for humans, but what about AIs?" https://www.theargumentmag.com/ ...
  • @deenamousa Deena Mousa on x
    Cool piece from @KelseyTuoc on whether LLMs are biased by the language the asker is using — mildly surprising to me that the answer was ‘no’! (h/t @otis_reid) [image]
  • @mattyglesias Matthew Yglesias on x
    Chatbots espousing cosmopolitan liberal values in all languages could have some interesting implications for social change https://substack.com/... [image]
  • @luke_metro @luke_metro on x
    Honestly it's kind of based that whoever is best at linear algebra is able to impose their values on the entire world