Amplifying Trans and Nonbinary Voices: A Community-Centred Harm Taxonomy for LLMs
Eddie L. Ungless, Sunipa Dev, Cynthia L. Bennett, Rebecca Gulotta, Jasmijn Bastings, Remi Denton
摘要
Warning: some of the example prompts given in this paper are offensive including slurs. These are intended to illustrate potential harms. We explore large language model (LLM) responses that may negatively impact the transgender and nonbinary (TGNB) community and introduce the Transing Transformers Toolkit, T 3 , which provides resources for identifying such harmful response behaviors. The heart of T 3 is a community-centred taxonomy of harms, developed in collaboration with the TGNB community, which we complement with, amongst other guidance, suggested heuristics for evaluation. To develop the taxonomy, we adopted a multi-method approach that included surveys and focus groups with community experts. The contribution highlights the importance of community-centred approaches in mitigating harm, and outlines pathways for LLM developers to improve how their models handle TGNB-related topics. * Work conducted as a student researcher at Google Research.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated StoriesKirti Bhagat, Shaily Bhatt, Athul Velagapudi, Aditya Vashistha 等CHI 2026 · 被引用 1 次
- Mind the Inclusivity Gap: Multilingual Gender-Neutral Translation Evaluation with mGeNTEBeatrice Savoldi, Giuseppe Attanasio, Eleonora Cupin, Eleni Gkovedarou 等EMNLP 2025
它引用的顶会 Paper10
- WildChat: 1M ChatGPT Interaction Logs in the WildWenting Zhao, Xiang Ren, Jack Hessel, Claire Cardie 等ICLR 2024 · 被引用 504 次
- Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language TechnologiesSunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian 等EMNLP 2021 · 被引用 113 次
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 被引用 68 次
- "They only care to show us the wheelchair": disability representation in text-to-image AI modelsKelly Avery Mack, Rida Qadri, Remi Denton, Shaun K. Kane 等CHI 2024 · 被引用 57 次
- WinoQueer: A Community-in-the-Loop Benchmark for Anti-LGBTQ+ Bias in Large Language ModelsVirginia K. Felkner, Ho-Chun Herbert Chang, Eugene Jang, Jonathan MayACL 2023 · 被引用 46 次
相关 Paper
- GPT is Not an Annotator: The Necessity of Human Annotation in Fairness Benchmark ConstructionVirginia K. Felkner, Jennifer A. Thompson, Jonathan MayACL 2024 · 被引用 3 次
- Trust The TypicalDebargha Ganguly, Sreehari Sankar, Biyao Zhang, Vikash Singh 等ICLR 2026 · 被引用 3 次
- SafetyKit: First Aid for Measuring Safety in Open-domain Conversational SystemsEmily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon L. Spruit 等ACL 2022
- Probing Toxic Content in Large Pre-Trained Language ModelsNedjma Ousidhoum, Xinran Zhao, Tianqing Fang, Yangqiu Song 等ACL 2021
- BingoGuard: LLM Content Moderation Tools with Risk LevelsFan Yin, Philippe Laban, Xiangyu Peng, Yilun Zhou 等ICLR 2025
