STL: Still Tricky Logic (for System Validation, Even When Showing Your Work)
Isabelle Hurley, Rohan Paleja, Ashley Suh, Jaime Daniel Peña, Ho Chit Siu
摘要
As learned control policies become increasingly common in autonomous systems, there is increasing need to ensure that they are interpretable and can be checked by human stakeholders. Formal specifications have been proposed as ways to produce human-interpretable policies for autonomous systems that can still be learned from examples. Previous work showed that despite claims of interpretability, humans are unable to use formal specifications presented in a variety of ways to validate even simple robot behaviors. This work uses active learning, a standard pedagogical method, to attempt to improve humans' ability to validate policies in signal temporal logic (STL). Results show that overall validation accuracy is not high, at (mean standard deviation), and that the three conditions of no active learning, active learning, and active learning with feedback do not significantly differ from each other. Our results suggest that the utility of formal specifications for human interpretability is still unsupported but point to other avenues of development which may enable improvements in system validation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-makingZana Buçinca, Maja Barbara Malaya, Krzysztof Z. GajosCSCW 2021 · 被引用 962 次
- Interpretable and Personalized Apprenticeship Scheduling: Learning Interpretable Scheduling Policies from Heterogeneous User DemonstrationsRohan R. Paleja, Andrew Silva, Letian Chen, Matthew C. GombolayNeurIPS 2020 · 被引用 41 次
相关 Paper
- Learning Branching-Time Properties in CTL and ATL via Constraint SolvingBenjamin Bordais, Daniel Neider, Rajarshi RoyFM 2024 · 被引用 4 次
- Automaton Constrained Q-LearningAnastasios Manganaris, Vittorio Giammarino, Ahmed H. QureshiNeurIPS 2025 · 被引用 3 次
- RESTL: Reinforcement Learning Guided by Multi-Aspect Rewards for Signal Temporal Logic TransformationYue Fang, Zhi Jin, Jie An, Hongshen Chen 等AAAI 2026 · 被引用 1 次
- SafeDec: Constrained Decoding for Safe Autoregressive Generalist Robot Navigation PoliciesParv Kapoor, Akila Ganlath, Michael Clifford, Changliu Liu 等ICML 2026 · 被引用 1 次
- Learning Reliable and Intuitive Temporal Logic Rules for Interpretable Time Series ClassificationYang Wang, Jiaqi Zhu, Miaomiao Li, Jiang Liu 等KDD 2025
