Lune

NeurIPS2024Top-tier venue

STL: Still Tricky Logic (for System Validation, Even When Showing Your Work)

Isabelle Hurley, Rohan Paleja, Ashley Suh, Jaime Daniel Peña, Ho Chit Siu

2024Year
9Citations

Abstract

As learned control policies become increasingly common in autonomous systems, there is increasing need to ensure that they are interpretable and can be checked by human stakeholders. Formal specifications have been proposed as ways to produce human-interpretable policies for autonomous systems that can still be learned from examples. Previous work showed that despite claims of interpretability, humans are unable to use formal specifications presented in a variety of ways to validate even simple robot behaviors. This work uses active learning, a standard pedagogical method, to attempt to improve humans' ability to validate policies in signal temporal logic (STL). Results show that overall validation accuracy is not high, at 65%±15%65\% \pm 15\% (mean ±\pm standard deviation), and that the three conditions of no active learning, active learning, and active learning with feedback do not significantly differ from each other. Our results suggest that the utility of formal specifications for human interpretability is still unsupported but point to other avenues of development which may enable improvements in system validation.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 435e8451-4032-4ef5-b7f3-d8e48284bb73

Builds on2

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines