Skip to content

Latest commit

 

History

13 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 
 
 
 
 
 
 
 
 

Repository files navigation

Social Bias in LLMs and VLMs

papers
This curated list brings together key methods, datasets, and benchmarks related to social bias in LLMs and VLMs.

As multimodal models continue to advance, understanding, evaluating, and mitigating social bias has become an increasingly important research direction. This repository is intended as a practical index for researchers and engineers to quickly track relevant work and recent progress.

Contributions are welcome—feel free to open a Pull Request to add high-quality resources.

Table of Contents

Foundations

Sociology

Title Introduction Date Code
Publish
Getting a Job: Is There a Motherhood Penalty?
Motherhood Penalty 2007-03 -
Publish
A Meta-Analytic Test of Intergroup Contact Theory
Contact Works 2006-05 -
Publish
Are Emily and Greg More Employable Than Lakisha and Jamal? A Field Experiment
The Quality Paradox 2004-09 Code
Publish
Measuring Individual Differences in Implicit Cognition: The Implicit Association Test
Latency is the Key 1998-06 -
Publish
Rethinking Racism: Toward a Structural Interpretation
Racialized Social System 1997-06 -
Publish
Opposition to Race-Targeting: Self-Interest, Stratification Ideology, or Racial Attitudes?
Self-Interest is weak 1993-08 -
Publish
Understanding Everyday Racism: An Interdisciplinary Theory
Racism is Routine 1991-07 -
Publish
Race Prejudice as a Sense of Group Position
Group-level “sense of Position” 1958-03 -
Publish
Measuring Social Distances
Degree of Intimacy 1925-03 -

NLP (Pre-LLM)

Title Introduction Date Code
Publish
Language (Technology) is Power: A Critical Survey of “Bias” in NLP
- 2020-07 -
Publish
Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image Representations
image 2019-10 Github

Survey

LLM Surveys

Title Introduction Date Code
Publish
Bias and Fairness in Large Language Models: A Survey
image 2024-09 Github
Publish
Fairness in Large Language Models: A Taxonomic Survey
image 2024-07 -
Publish
An Empirical Survey of the Effectiveness of Debiasing Techniques for Pre-trained Language Models
- 2022-05 Github

VLM Surveys

Title Introduction Date Code
Publish
Fairness in Deep Learning: A Survey on Vision and Language Research
image 2025-02 -
Survey of Social Bias in Vision-Language Models image 2023-09 -

Data & Benchmarks

LLMs

Title Introduction Date Code
Publish
Performance unfairness of large language models in cross-language fact-checking
image 2026-06 -
How Quantization Shapes Bias in Large Language Models image 2026-01 Github
Publish
What’s Not Said Still Hurts: A Description-Based Evaluation Framework for Measuring Social Bias in LLMs
image 2025-09 Github
Publish
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
image 2025-08 Github
Publish
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
image 2025-04 -
Publish
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
image 2025-04 Github
Publish
Evaluating and Mitigating Discrimination in Language Model Decisions
image 2024-12 -
Publish
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
image 2024-10 Github
Publish
Ask LLMs Directly, “What shapes your bias? ”: Measuring Social Bias in Large Language Models
image 2024-08 Github(BBQ)
Publish
CHBias: Bias Evaluation and Mitigation of Chinese Conversational Language Models
image 2023-05 Github
Publish
BBQ: A Hand-Built Bias Benchmark for Question Answering
image 2022-05 Github

VLMs

Title Introduction Date Code
BBQ-V: Benchmarking Visual Stereotype Bias in Large Multimodal Models image 2026-01 Github
Publish
Vision Language Models are Biased
image 2026-01 Code
Publish
VISBIAS: Measuring Explicit and Implicit Social Biases in Vision Language Models
image 2025-11 Github
Publish
Evaluating Fairness in Large Vision-Language Models Across Diverse Demographic Attributes and Prompts
image 2025-11 -

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
image 2025-05 Github
HumaniBench: A Human-Centric Framework for Large Multimodal Models Evaluation image 2025-05 Github
Publish
MultiTrust: A Comprehensive Benchmark Towards Trustworthy Multimodal Large Language Models
image 2024-12 Code
Publish
ModSCAN: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
image 2024-11 Github
FMBench: Benchmarking Fairness in Multimodal Large Language Models on Medical Tasks image 2024-10 -
Publish
Are Gender-Neutral Queries Really Gender-Neutral? Mitigating Gender Bias in Image Search
image 2021-11 Github

Mitigation

Title Introduction Date Code
Observations and Remedies for Large Language Model Bias in Self-Consuming Performative Loop image 2026-01 -
Publish
Open-DeBias: Toward Mitigating Open-Set Bias in Language Models
image 2025-11 Resource
Debiasing Reward Models by Representation Learning with Guarantees image 2025-10 -
Publish
Robustly Improving LLM Fairness in Realistic Settings via Interpretability
image 2025-09 Github
Publish
Guiding LLM Decision-Making with Fairness Reward Models
image 2025-09 Github
Publish
Wisdom from Diversity: Bias Mitigation Through Hybrid Human-LLM Crowds
image 2025-08 -
Publish
Joint Vision-Language Social Bias Removal for CLIP
image 2025-06 Github
Publish
Prompt Tuning Pushes Farther, Contrastive Learning Pulls Closer: A Two-Stage Approach to Mitigate Social Biases
image 2023-07 -
Publish
BLIND: Bias Removal With No Demographics
image 2023-07 Github
Publish
Don’t Just Clean It, Proxy Clean It: Mitigating Bias by Proxy in Pre-Trained Models
image 2022-12 Data:
WIKI
Madlibs
BIOS
Publish
Fairness without Demographics through Adversarially Reweighted Learning
image 2020-11 -

Findings

Title Introduction Date Code
Biases in the Blind Spot: Detecting What LLMs Fail to Mention image 2026-02 -
FairReason: Balancing Reasoning and Social Bias in MLLMs image 2025-09 Github
Publish
Evaluating Sex and Age Biases in Multimodal Large Language Models for Skin Disease Identification from Dermatoscopic Images
image 2025-04 -
Publish
Systematic Biases in LLM Simulations of Debates
image 2024-12 -
Publish
Large Language Models are not Fair Evaluators
image 2024-07 Github
Publish
Data Feedback Loops: Model-driven Amplification of Dataset Biases
image 2023-07 Github
Publish
Upstream Mitigation Is Not All You Need: Testing the Bias Transfer Hypothesis in Pre-Trained Language Models
image 2022-05 Adapted code:
HurtfulWords
sent_debias
biosbias
unintended-ml-bias-analysis
roberta-base
Publish
Challenges in Automated Debiasing for Toxic Language Detection
image 2021-04 Github

Disclaimer

This repository is intended for educational and research purposes only. The images and figures displayed in this repository are the property of their respective authors and publishers and are used here for illustrative purposes to facilitate quick browsing and understanding of the papers.

If you are a copyright holder and would like any content removed, please open an issue, and I will remove it promptly.

About

A paper list for social bias in LLMs and VLMs.

Topics

Resources

Stars

8 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages