Do frontier LLMs represent how national cultures change?
Evaluated Claude, GPT, Gemini and Qwen against two decades of World Values Survey data across 40 countries.
Quantitative behavioral scientist working on AI evaluation, trust in AI, and multilingual & cultural safety. My focus is measurement: operationalizing constructs like trust, bias, harm and cultural alignment so evaluations capture what they claim to.
Whether LLMs represent people outside the Western, English-speaking world, how they are reshaping cultures, and where safety fails in other languages.
How people come to trust AI systems, when that trust is warranted, and how relying on them changes decisions.
The sociotechnical and ethical consequences of living with these systems, including how interacting with AI shapes people's sense of self.
How moral values shift over time and differ across societies, and how groups misread each other's values (the psychology the AI work is built on).
Evaluated Claude, GPT, Gemini and Qwen against two decades of World Values Survey data across 40 countries.
Examining how AI systems may standardize cultural representations and what this means for policy and cultural diversity.
10M+ tweets and three preregistered experiments show a low-cost correction raises intergroup trust.
Coverage of "The Homogenizing Engine" on how AI systems can flatten cultural differences, and what policy can do about it.
Read the story →Coverage of "The Homogenizing Engine" and its findings on whose worldview AI systems reproduce.
Read the story →SPSP featured and published the science advocacy field guide I wrote after the AAAS CASE Workshop.
Read the story →