The Promises and Perils of using LLMs for Effective Public Services
Erina Seh-Young Moon, Matthew Tamura, Angelina Zhai, Nuzaira Habib, Behnaz Shirazi, Altaf Kassam, Devansh Saxena, Shion Guha
Abstract
Governments are the primary providers of essential public services and are responsible for delivering them effectively. In high-stakes decision-making domains such as child welfare (CW), agencies must protect children without unnecessarily prolonging a family’s engagement with the system. With growing optimism around AI, governments are pushing for its integration but concerns regarding feasibility and harms remain. Through collaborations with a large Canadian CW agency, we examined how LocalLLM and BERTopic models can track CW case progress. We demonstrate how the tools can potentially assist workers in opportunistically addressing gaps in their work by signaling case progress/deviations. And yet, we also show how they fail to detect case trajectories that require discretionary judgments grounded in social work training, areas where practitioners would actually want support to pre-emptively address substantive case concerns. We also provide a roadmap of future participatory directions to co-design language tools for/with the public sector.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on38
- Critical Race Theory for HCIIhudiya Finda Ogbonnaya-Ogburu, Angela D. R. Smith, Alexandra To, Kentaro ToyamaCHI 2020 · 397 citations
- A Case for Humans-in-the-Loop: Decisions in the Presence of Erroneous Algorithmic ScoresMaria De-Arteaga, Riccardo Fogliato, Alexandra ChouldechovaCHI 2020 · 176 citations
- Improving Human-AI Partnerships in Child Welfare: Understanding Worker Practices, Challenges, and Desires for Algorithmic Decision SupportAnna Kawakami, Venkatesh Sivaraman, Hao Fei Cheng, Logan Stapleton et al.CHI 2022 · 137 citations
- Jury Learning: Integrating Dissenting Voices into Machine Learning ModelsMitchell L. Gordon, Michelle S. Lam, Joon Sung Park, Kayur Patel et al.CHI 2022 · 134 citations
- A Framework of High-Stakes Algorithmic Decision-Making for the Public Sector Developed through a Case Study of Child-WelfareDevansh Saxena, Karla A. Badillo-Urquiola, Pamela J. Wisniewski, Shion GuhaCSCW 2021 · 133 citations
Related papers
- Automate, Assist, Avoid: Caseworkers' Perspectives on Applying Large Language Model-Based Assistance in Public Sector Decision-Making ProcessesKarolina Drobotowicz, Johanna Ylipulli, Uttishta Sreerama Varanasi, Heidi S. MäkitaloCHI 2026 · 1 citation
- Understanding Public Agencies' Expectations and Realities of AI-Driven Chatbots for Public Health MonitoringEunkyung Jo, Young-Ho Kim, Sang-Houn Ok, Daniel A. EpsteinCHI 2025 · 11 citations
- The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder, Early-stage Deliberations Around Public Sector AI ProposalsAnna Kawakami, Amanda Coston, Haiyi Zhu, Hoda Heidari et al.CHI 2024 · 54 citations
- Addressing Procedural and Tooling Challenges in Juvenile Justice: Towards Responsible and Human-Centered DesignElizabeth S. Gilman, Christopher Flathmann, Emma DixonCHI 2026 · 1 citation
- Towards Aligning Multimodal LLMs with Human Experts: A Focus on Parent-Child InteractionWeiyan Shi, Kenny Tsu Wei ChooCHI 2026 · 2 citations
