arrow
Return

Implicit-bias-like patterns in reasoning models

delete2026-09-01
delete0
PRE
AI
M
Messi Ho Jun Lee *
C
Calvin K. Lai
DOI:10.1038/s42256-026-01300-1delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Implicit biases refer to automatic mental processes that shape perceptions, judgements and behaviours. While previous research on bias in large language models (LLMs) has focused primarily on outputs, we introduce the reasoning-model implicit association test (RM-IAT) to study bias-like processing differences in reasoning models (that is, LLMs that generate explicit step-by-step reasoning before producing a response). By measuring reasoning-token counts as an index of computational effort, the RM-IAT captures processing efficiency differences analogous to response latency differences in the human IAT. Across four models (o3-mini, DeepSeek-R1, gpt-oss-20b and Qwen3-8B), we find consistent evidence that association-incompatible tasks require greater computational effort than association-compatible tasks. Claude 3.7 Sonnet exhibited reversed patterns, which were linked using thematic analysis to its unique internal focus on reasoning about bias and stereotypes. We also found evidence for convergent validity with model outputs. RM-IAT effects predicted biases in two tasks known to capture LLM biases in word association and decision-making. Together, these findings demonstrate that the RM-IAT captures meaningful variation in how reasoning models process stereotypical information, and that this variation predicts downstream model behaviour. Lee and Lai study bias-like processing differences in large language reasoning models and find that, for most models, processing stereotypical information takes less computational effort than processing counter-stereotypical information.

Journal

Nature Machine Intelligence cover
Nature Machine Intelligence
IF:
23.9
Papers:
1.3K
Citations:
1.5W

Organization

W
Washington University in St. Louis
Scholars:
1.0K
Papers: 425
Citations: 0
R
rutgers university
Scholars:
1.4K
Papers: 821
Citations: 0
Cited Papers

Cited Papers

errShare
errSave
Implicit Social Cognition
err2020-01-04
err218
PREAI
errGreenwald, Anthony G.; Lai, Calvin K.
errShare
errSave
errShare
errSave
Implicit Bias and Policing
err2016-01-08
err0
PREAI
errKatherine B. Spencer; Amanda K. Charbonneau; Jack Glaser
errShare
errSave
The Mythical Number Two
err2018-04-01
err247
PREAI
errMelnikoff, David E.; Bargh, John A.
errShare
errSave
Fitting Linear Mixed-Effects Models Using lme4
err2015-01-01
err6.0W
errOAAI
errBates, Douglas; Maechler, Martin; Bolker, Benjamin M.; Walker, Steven C.
errShare
errSave
researcher View more