Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)

Leander Girrbach; Stephan Alaniz; Yiran Huang; Trevor Darrell; Zeynep Akata

Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)

Leander Girrbach, Stephan Alaniz, Yiran Huang, Trevor Darrell, Zeynep Akata

TL;DR

The paper addresses gender bias in vision-language assistants by introducing VL-Gender, a framework that evaluates biases across personality traits, skills, and occupations using 22 open-source VLAs and a curated, occupation-free image dataset. It employs robust prompt variations and a VQA-style evaluation to reveal biases that often mirror real-world gender imbalances, such as male associations with negative traits and female associations with positive traits, with some models also displaying real-world occupation biases. Debiasing experiments compare five methods, finding that full fine-tuning offers the strongest bias reduction with acceptable performance costs, while other methods provide more conservative or task-preserving adjustments. The work emphasizes pre-deployment bias assessment, reproducibility, and the need for scalable debiasing strategies to promote equitable societal outcomes in VLAs.

Abstract

Pre-trained large language models (LLMs) have been reliably integrated with visual input for multimodal tasks. The widespread adoption of instruction-tuned image-to-text vision-language assistants (VLAs) like LLaVA and InternVL necessitates evaluating gender biases. We study gender bias in 22 popular open-source VLAs with respect to personality traits, skills, and occupations. Our results show that VLAs replicate human biases likely present in the data, such as real-world occupational imbalances. Similarly, they tend to attribute more skills and positive personality traits to women than to men, and we see a consistent tendency to associate negative personality traits with men. To eliminate the gender bias in these models, we find that fine-tuning-based debiasing methods achieve the best trade-off between debiasing and retaining performance on downstream tasks. We argue for pre-deploying gender bias assessment in VLAs and motivate further development of debiasing strategies to ensure equitable societal outcomes. Code is available at https://github.com/ExplainableML/vla-gender-bias.

Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)

TL;DR

Abstract

Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)

TL;DR

Abstract

Paper Structure

Table of Contents

Figures (25)