← All news

Safety

Study finds tested language models formed hiring patterns more stratified than a human baseline

A study using a simulated, repeated hiring task found that tested large language models developed group-to-job patterns more stratified than a human baseline, despite equal underlying success probabilities for the artificial groups.

reviewedUpdated Aug 14, 2026, 7:36 PM UTC
Original source

arXiv

Read the original source

What happened

Researchers tested large language models in a multi-round hiring game adapted from a psychology experiment. The task used invented demographic groups, stylized jobs, repeated feedback, and equal underlying chances of success.

The paper reports that the models learned to assign artificial groups to different jobs more strongly than the study’s human-participant baseline. An independent report described tests involving ChatGPT, Claude, and Gemini.

The result concerns generalization from sequential feedback, not production résumé screening. It does not measure real applicants, employment decisions, or discrimination in deployed hiring systems.

Why it matters

Hiring tools can influence who advances before a human review. The research suggests that bias risks may arise not only from patterns in training data, but also from how models adapt to feedback during repeated decisions. The study does not establish that this effect occurs in real hiring.

What remains unclear

Related claims

Sources

  1. Large Language Models Develop Novel Social Biases Through Adaptive ExplorationPrimary source - arXiv - research preprint / icml paper record - Nov 8, 2025

    Used for: Study design, findings, and limitations.

    Open source

  2. Princeton and Chicago: LLMs stereotype hires more than humansAI Weekly - secondary news analysis - Jul 27, 2026

    Used for: Independent description of the simulated hiring experiment and its scope.

    Open source