Amplifying Trans and Nonbinary Voices: A Community-Centred Harm Taxonomy for LLMs
Eddie L. Ungless, Sunipa Dev, Cynthia L. Bennett, Rebecca Gulotta, Jasmijn Bastings, Remi Denton
Abstract
Warning: some of the example prompts given in this paper are offensive including slurs. These are intended to illustrate potential harms. We explore large language model (LLM) responses that may negatively impact the transgender and nonbinary (TGNB) community and introduce the Transing Transformers Toolkit, T 3 , which provides resources for identifying such harmful response behaviors. The heart of T 3 is a community-centred taxonomy of harms, developed in collaboration with the TGNB community, which we complement with, amongst other guidance, suggested heuristics for evaluation. To develop the taxonomy, we adopted a multi-method approach that included surveys and focus groups with community experts. The contribution highlights the importance of community-centred approaches in mitigating harm, and outlines pathways for LLM developers to improve how their models handle TGNB-related topics. * Work conducted as a student researcher at Google Research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 141c7b47-6423-4a39-acb0-c98e4270dcebCited by top-tier papers2
- TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated StoriesKirti Bhagat, Shaily Bhatt, Athul Velagapudi, Aditya Vashistha et al.CHI 2026 · 1 citation
- Mind the Inclusivity Gap: Multilingual Gender-Neutral Translation Evaluation with mGeNTEBeatrice Savoldi, Giuseppe Attanasio, Eleonora Cupin, Eleni Gkovedarou et al.EMNLP 2025
Builds on10
- WildChat: 1M ChatGPT Interaction Logs in the WildWenting Zhao, Xiang Ren, Jack Hessel, Claire Cardie et al.ICLR 2024 · 504 citations
- Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language TechnologiesSunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian et al.EMNLP 2021 · 113 citations
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 68 citations
- "They only care to show us the wheelchair": disability representation in text-to-image AI modelsKelly Avery Mack, Rida Qadri, Remi Denton, Shaun K. Kane et al.CHI 2024 · 57 citations
- WinoQueer: A Community-in-the-Loop Benchmark for Anti-LGBTQ+ Bias in Large Language ModelsVirginia K. Felkner, Ho-Chun Herbert Chang, Eugene Jang, Jonathan MayACL 2023 · 46 citations
Related papers
- GPT is Not an Annotator: The Necessity of Human Annotation in Fairness Benchmark ConstructionVirginia K. Felkner, Jennifer A. Thompson, Jonathan MayACL 2024 · 3 citations
- Trust The TypicalDebargha Ganguly, Sreehari Sankar, Biyao Zhang, Vikash Singh et al.ICLR 2026 · 3 citations
- SafetyKit: First Aid for Measuring Safety in Open-domain Conversational SystemsEmily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon L. Spruit et al.ACL 2022
- Probing Toxic Content in Large Pre-Trained Language ModelsNedjma Ousidhoum, Xinran Zhao, Tianqing Fang, Yangqiu Song et al.ACL 2021
- BingoGuard: LLM Content Moderation Tools with Risk LevelsFan Yin, Philippe Laban, Xiangyu Peng, Yilun Zhou et al.ICLR 2025
