PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation
Yu Yang, Ig-Jae Kim, Dongwook Yoon
摘要
AI compliance is becoming increasingly critical as AI systems grow more powerful and pervasive. Yet the rapid expansion of AI policies creates substantial burdens for resource-constrained practitioners lacking policy expertise. Existing approaches typically address one policy at a time, making multi-policy compliance costly. We present PASTA, a scalable compliance tool integrating four innovations: (1) a comprehensive model-card format supporting descriptive inputs across development stages; (2) a policy normalization scheme ; (3) an efficient LLM-powered pairwise evaluation engine with cost-saving strategies; and (4) an interface delivering interpretable evaluations via compliance heatmaps and actionable recommendations. Expert evaluation shows PASTA's judgments closely align with human experts (𝜌 ≥ .626). The system evaluates five major policies in under two minutes at approximately $3. A user study (N = 12) confirms practitioners found outputs easy-to-understand and actionable, introducing a novel framework for scalable automated AI governance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Farsight: Fostering Responsible AI Awareness During AI Application PrototypingZijie J. Wang, Chinmay Kulkarni, Lauren Wilcox, Michael Terry 等CHI 2024 · 被引用 55 次
- The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder, Early-stage Deliberations Around Public Sector AI ProposalsAnna Kawakami, Amanda Coston, Haiyi Zhu, Hoda Heidari 等CHI 2024 · 被引用 54 次
- Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit ToolingVictor Ojewale, Ryan Steed, Briana Vecchione, Abeba Birhane 等CHI 2025 · 被引用 46 次
- RAI Guidelines: Method for Generating Responsible AI Guidelines Grounded in Regulations and Usable by (Non-)Technical RolesMarios Constantinides, Edyta Paulina Bogucka, Daniele Quercia, Susanna Kallio 等CSCW 2024 · 被引用 28 次
- RiskRAG: A Data-Driven Solution for Improved AI Model Risk ReportingPooja S. B. Rao, Sanja Scepanovic, Ke Zhou, Edyta Paulina Bogucka 等CHI 2025 · 被引用 7 次
相关 Paper
- COMPASS: A Framework for Evaluating Organization-Specific Policy Alignment in LLMsDasol Choi, DongGeon Lee, Brigitta Jesica Kartono, Helena Berndt 等ACL 2026 · 被引用 3 次
- PolicyPad: Collaborative Prototyping of LLM PoliciesK. J. Kevin Feng, Tzu-Sheng Kuo, Quan Ze Jim Chen, Inyoung Cheong 等CHI 2026 · 被引用 2 次
- Licoeval: Evaluating LLMs on License Compliance in Code GenerationWeiwei Xu, Kai Gao, Hao He, Minghui ZhouICSE 2025 · 被引用 6 次
- Privy: Envisioning and Mitigating Privacy Risks for Consumer-facing AI Product ConceptsHao-Ping (Hank) Lee, Yu-Ju Yang, Matthew Bilik, Isadora Krsek 等CHI 2026 · 被引用 1 次
- LLM Comparator: Interactive Analysis of Side-by-Side Evaluation of Large Language ModelsMinsuk Kahng, Ian Tenney, Mahima Pushkarna, Michael Xieyang Liu 等IEEE VIS 2024 · 被引用 23 次
