SAFO: Stable Adaptive Fairness Optimization for LLM-Based Social Survey Simulation
Chenxi Lin, Zhuoren Jiang, Kaisong Song, Yiquan Wu
摘要
Ensuring fairness in social survey simulation is critical, as biased outputs can misrepresent underrepresented groups. This issue is growing as large language models (LLMs) are increasingly used for this task. However, standard finetuning based on Empirical Risk Minimization (ERM) often under-optimizes minority groups, causing substantial subgroup disparities. Distributionally robust Optimization (DRO) methods reduce worst-case errors, but their strict worst-case selection can lead to noisy and unstable optimization under demographic sparsity. These issues create intertwined challenges for fairness, convergence and stability. We propose SAFO, a dynamic utility-fairness optimization framework for LLM-based survey simulation that explicitly targets both fairness and training stability. SAFO combines (i) an Optimizer that preserves mean-loss utility, (ii) an Adversary that performs temperature-controlled, EMAsmoothed and loss-driven group reweighting, and (iii) a Nash-inspired Regulator that adaptively adjusts the utility-fairness trade-off by tracking weak-group gains and collateral utility damages. Experiments on three large-scale survey datasets from China, the U.S., and Europe show that SAFO consistently improves minority performance and social-welfare metrics. It reduces worst-group gaps by up to 12.7%, maintains overall accuracy with a mean change of less than 0.3% and lowers variance across random seeds. Our code is available at https://github.com/PiLab-ZJU/SAFO .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Fair Influence Maximization: a Welfare Optimization ApproachAida Rahmattalabi, Shahin Jabbari, Himabindu Lakkaraju, Phebe Vayanos 等AAAI 2021 · 被引用 71 次
- Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public OpinionsJoseph Suh, Erfan Jahanparast, Suhong Moon, Minwoo Kang 等ACL 2025 · 被引用 48 次
- Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMsAngelina Wang, Michelle Phan, Daniel E. Ho, Sanmi KoyejoACL 2025 · 被引用 17 次
- Distributionally Robust Optimization with Probabilistic GroupSoumya Suvra Ghosal, Yixuan LiAAAI 2023 · 被引用 14 次
- On the Fairness ROAD: Robust Optimization for Adversarial DebiasingVincent Grari, Thibault Laugel, Tatsunori Hashimoto, Sylvain Lamprier 等ICLR 2024 · 被引用 6 次
相关 Paper
- Re-weighting Based Group Fairness Regularization via Classwise Robust OptimizationSangwon Jung, Taeeon Park, Sanghyuk Chun, Taesup MoonICLR 2023 · 被引用 5 次
- Does LLM Focus on the Right Words? Mitigating Context Bias in LLM-based RecommendersBohao Wang, Jiawei Chen, Feng Liu, Changwang Zhang 等WWW 2026 · 被引用 1 次
- MaxMin-RLHF: Alignment with Diverse Human PreferencesSouradip Chakraborty, Jiahao Qiu, Hui Yuan, Alec Koppel 等ICML 2024 · 被引用 104 次
- Task-level Distributionally Robust Optimization for Large Language Model-based Dense RetrievalGuangyuan Ma, Yongliang Ma, Xing Wu, Zhenpeng Su 等AAAI 2025 · 被引用 6 次
- Distributionally Robust Reinforcement Learning from Human FeedbackDebmalya Mandal, Paulius Sasnauskas, Goran RadanovicICML 2026
