PrometheusRoot
Blog Links Prometheans 100+ AI Books AI Companies Why are you here?
← Prometheans 100+
×
Yejin Choi
pioneer
Researcher
X / Twitter Website Wikipedia
common-sensereasoningmacarthur

Recognition

TIME 100 AI 2023 TIME 100 AI 2025

Related

builder Francois Chollet builder Gary Marcus
← Prometheans 100+ Yejin Choi
TIME 100 AI 2023 TIME 100 AI 2025

Stanford HAI Professor, commonsense AI researcher

Yejin Choi

Dieter Schwarz Foundation HAI Professor, Senior Fellow at HAI — Stanford University Professor (former) — University of Washington Researcher (former) — Allen Institute for AI
Listen — profile
0:00 / 3:04

Profile

Yejin Choi is the researcher who keeps asking the question the rest of the field would rather skip past: does a large language model actually understand anything, or is it an extraordinarily gifted mimic? As the Dieter Schwarz Foundation HAI Professor of Computer Science and a Senior Fellow at Stanford HAI, Choi has spent her career probing the gap between what AI systems can say and what they actually know — and she has built the benchmarks, datasets, and demos that make that gap impossible to ignore. For developers weaned on the “just scale it” narrative, her work is the necessary counterweight: proof that fluency and comprehension are not the same thing.

Choi’s path is unusual. Before academia she spent years as a software engineer, and she has spoken candidly about arriving at NLP research later and more circuitously than her peers — which may be why her instincts run against fashion. She built her reputation over a decade at the University of Washington and the Allen Institute for AI (AI2), where she led some of the most cited work in commonsense reasoning. In 2024 she left Seattle for California, joining NVIDIA as a Senior Director of AI Research before landing at Stanford in early 2025. Along the way she collected a 2022 MacArthur “genius” Fellowship — awarded specifically for teaching machines common sense — and appeared on TIME’s 100 Most Influential People in AI in both 2023 and 2025.

Her fingerprints are on much of the infrastructure the field uses to measure reasoning. ATOMIC and COMET turned commonsense knowledge into something a neural network could generate rather than just look up. Benchmarks like HellaSwag and WinoGrande were designed to be adversarially hard for models that were quietly cheating on easier tests — and they became standard yardsticks precisely because they exposed shortcuts. Grover demonstrated that the best detector of AI-generated fake news was the same kind of model that produced it. And Delphi, her most provocative project, asked whether a system could learn to make moral judgments at all. Each was less a product than an argument, delivered in code.

What makes Choi matter to anyone building with AI today is her intellectual honesty about limits. She is not a doomer and not a hype merchant; she’s an empiricist who cheerfully shows you a state-of-the-art model failing at a puzzle a six-year-old would solve. Like Gary Marcus, she doubts that scale alone delivers understanding — but unlike him, she spends her time constructing the experiments and smaller, norm-trained systems that could actually close the gap. In a research culture increasingly organized around parameter counts, Choi insists on asking what the numbers are actually measuring. That skepticism, grounded in real benchmarks, is what keeps the field honest.

Key Articles & Papers

COMET: Commonsense Transformers for Automatic Knowledge Graph Construction 2019 — Reframed commonsense knowledge as something a transformer could generate on the fly, not just retrieve — a foundational idea for reasoning systems. Defending Against Neural Fake News (Grover) 2019 — Showed that the strongest detector of machine-generated disinformation is a model of the same kind that produced it. HellaSwag: Can a Machine Really Finish Your Sentence? 2019 — An adversarially-filtered benchmark that became a standard test for commonsense sentence completion — and exposed how much models were shortcut-learning. WinoGrande: An Adversarial Winograd Schema Challenge at Scale 2019 — Scaled up the classic Winograd pronoun test and stripped out the biases models were exploiting, raising the bar for genuine reasoning. (COMET-)ATOMIC 2020: On Symbolic and Neural Commonsense Knowledge Graphs 2020 — Extended ATOMIC into a broad knowledge graph of everyday cause-and-effect, powering neural commonsense inference. Can Machines Learn Morality? The Delphi Experiment 2021 — A deliberately provocative attempt to model human moral judgments — and a case study in how hard, and contested, that goal is. The Curious Case of Commonsense Intelligence (Daedalus) 2022 — Her accessible manifesto on why common sense is the missing dark matter of AI.

Videos

YouTube video

Controversies

Choi’s Delphi demo (2021) drew the sharpest reaction of her career. Released as an interactive tool that would render moral verdicts on user-typed scenarios, it quickly produced offensive and inconsistent judgments when probed adversarially, and critics — in the press and in academia — argued that framing ethics as a data-labeling classification problem was both technically shaky and philosophically naive. Choi’s team responded by relabeling Delphi as a research prototype for modeling people’s moral judgments rather than an oracle of right and wrong, and later published a peer-reviewed account in Nature Machine Intelligence. Whether one reads the episode as overreach or as valuable stress-testing of a genuinely hard question, it remains the clearest example of Choi pushing an idea far enough to draw real fire — which is arguably the point of the work.

Spotify Podcasts

The Evolution of Reasoning in Small Language Models with Yejin Choi - #761
The Evolution of Reasoning in Small Language Models with Yejin Choi - #761
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
2026
Episode 5: Yejin Choi
Episode 5: Yejin Choi
Unconfuse Me with Bill Gates
2023
Yejin Choi - Natural Language Processing, Common Sense, AI • YASP #4
Yejin Choi - Natural Language Processing, Common Sense, AI • YASP #4
Yet Another Science Podcast
2023
248 | Yejin Choi on AI and Common Sense
248 | Yejin Choi on AI and Common Sense
Sean Carroll's Mindscape: Science, Society, Philosophy, Culture, Arts, and Ideas
2023
Episode 248 | Yejin Choi on AI and Common Sense
Episode 248 | Yejin Choi on AI and Common Sense
Sean Carroll's Mindscape
2023
Yejin Choi: teaching AI common sense and morality
Yejin Choi: teaching AI common sense and morality
The Robot Brains Podcast
2023
Why AI is incredibly smart -- and shockingly stupid | Yejin Choi
Why AI is incredibly smart -- and shockingly stupid | Yejin Choi
TED Talks Daily
2023
Yejin Choi: Teaching Machines Common Sense and Morality
Yejin Choi: Teaching Machines Common Sense and Morality
The Gradient: Perspectives on AI
2022
S2E20: Yejin Choi with Dhruv Batra
S2E20: Yejin Choi with Dhruv Batra
Humans of AI: Stories, Not Stats
2021
Social Commonsense Reasoning with Yejin Choi - #518
Social Commonsense Reasoning with Yejin Choi - #518
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
2021

YouTube

YouTube video
2026
YouTube video
2024
YouTube video
2023
YouTube video
2023
YouTube video
2023
YouTube video
2023
YouTube video
2023
YouTube video
2023
YouTube video
2023
YouTube video
2021

Related People

builder Francois Chollet builder Gary Marcus
© 2026 PrometheusRoot