Job Unfair: An Investigation of Gender and Occupational Bias in Free-Form Text Completions by LLMs
Camilla Casula, Sebastiano Vecellio Salto, Elisa Leonardelli, Sara Tonelli
Abstract
Disentangling how gender and occupations are encoded by LLMs is crucial to identify possible biases and prevent harms, especially given the widespread use of LLMs in sensitive domains such as human resources. In this work, we carry out an in-depth investigation of gender and occupational biases in English and Italian as expressed by 9 different LLMs (both base and instruction-tuned). Specifically, we focus on the analysis of sentence completions when LLMs are prompted with job-related sentences including different gender representations. We carry out a manual analysis of 4,500 generated texts over 4 dimensions that can reflect bias, we propose a novel embedding-based method to investigate biases in generated texts and, finally, we carry out a lexical analysis of the model completions. In our qualitative and quantitative evaluation we show that many facets of social bias remain unaccounted for even in aligned models, and LLMs in general still reflect existing gender biases in both languages. Finally, we find that models still struggle with genderneutral expressions, especially beyond English. Model Subject M (%) F (%) N (%) aya-expanse-8b Abstract 0.00% 0.00% 3.66% Object 0.00% 0.00% 1.22% Profession 0.00% 0.00% 0.00% gemma-7b Abstract 0.00% 0.00% 9.09% Object 0.00% 1.15% 2.27% Profession 0.00% 0.00% 0.00% gemma-7b-instruct Abstract 0.00% 0.00% 1.22% Object 0.00% 0.00% 0.00% Profession 0.00% 0.00% 0.00% Llama-70B Abstract 0.00% 0.00% 1.47% Object 0.00% 0.00% 1.47% Profession 0.00% 0.00% 1.47% Llama-70B-instruct Abstract 0.00% 0.00% 2.56% Object 0.00% 0.00% 0.00% Profession 1.22% 0.00% 1.28% Llama-8B Abstract 0.00% 0.00% 1.18% Object 0.00% 0.00% 2.35% Profession 0.00% 0.00% 1.18% Llama-8B-instruct Abstract 0.00% 0.00% 0.00% Object 0.00% 0.00% 0.00% Profession 0.00% 0.00% 0.00% Mistral-7B Abstract 0.00% 0.00% 3.33% Object 0.00% 0.00% 1.11% Profession 0.00% 0.00% 1.11% Mistral-7B-instruct Abstract 0.00% 0.00% 1.16% Object 0.00% 0.00% 1.16% Profession 5.62% 3.80% 4.65%
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4461ce25-cbdb-4edf-88de-b844b70eb46bCited by top-tier papers1
Ask how each one uses itBuilds on9
- Calibrate Before Use: Improving Few-shot Performance of Language ModelsZihao Zhao, Eric Wallace, Shi Feng, Dan Klein et al.ICML 2021 · 1,843 citations
- Large Language Models Are Not Robust Multiple Choice SelectorsChujie Zheng, Hao Zhou, Fandong Meng, Jie Zhou et al.ICLR 2024 · 424 citations
- Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language TechnologiesSunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian et al.EMNLP 2021 · 113 citations
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 68 citations
- CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language ModelsNikita Nangia, Clara Vania, Rasika Bhalerao, Samuel R. BowmanEMNLP 2020 · 19 citations
Related papers
- On the Mutual Influence of Gender and Occupation in LLM RepresentationsHaozhe An, Connor Baumler, Abhilasha Sancheti, Rachel RudingerACL 2025
- Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)Leander Girrbach, Stephan Alaniz, Yiran Huang, Trevor Darrell et al.ICLR 2025
- GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language ModelsTao Zhang, Ziqian Zeng, Yuxiang Xiao, Huiping Zhuang et al.ACL 2025 · 18 citations
- Angry Men, Sad Women: Large Language Models Reflect Gendered Stereotypes in Emotion AttributionFlor Miriam Plaza del Arco, Amanda Cercas Curry, Alba Cercas Curry, Gavin Abercrombie et al.ACL 2024 · 8 citations
- Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji EmbeddingsMingchen Li, Wajdi Aljedaani, Yingjie Liu, Navyasri Meka et al.WWW 2026
