Researchers at Princeton University and the University of Chicago published a study that found large language models, including ChatGPT, Claude, and Gemini, learned hiring-related stereotypes more aggressively than human participants in simulated hiring experiments, with higher-reasoning models showing the strongest segregation.