Gender bias and stereotypes in Large Language Models
Large Language Models (LLMs) have made substantial progress in the past several months, shattering state-of-the-art benchmarks in many domains. This paper investigates LLMs’ behavior with respect to gender stereotypes, a known issue for prior models. We use a simple paradigm to test the presence of gender bias, building on but differing from WinoBias, a commonly used gender bias dataset, which is likely to be included in the training data of current LLMs. We test four recently published LLMs and
Record details
Published: 13 October 2023
Source: OpenAlex
Category: Research
Topics: Bias & fairness · Finance, VC & PE
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Journal of Management World
OpenAlex · 31 January 2024
Comparing Generative AI and teacher feedback: student perceptions of usefulness and trustworthiness
OpenAlex · 13 May 2025
The Geometric Inductive Bias of Grokking: Bypassing Phase Transitions via Architectural Topology
arXiv · 5 March 2026
Analysis Of Linguistic Stereotypes in Single and Multi-Agent Generative AI Architectures
arXiv · 19 March 2026
Who Benefits from RAG? The Role of Exposure, Utility and Attribution Bias
arXiv · 25 March 2026
QoS-Aware Token Scheduling and Private Data Valuation for Multi-Modal Agentic Networks
arXiv · 2 April 2026
How to cite this record
ethics.ai (13 October 2023), “Gender bias and stereotypes in Large Language Models,” evidence record 9432, https://ethics.ai/record/9432 (originally published by OpenAlex).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.