On Detecting Cherry-picked Generalizations
Yin Lin, Brit Youngmann, Yuval Moskovitch, H. V. Jagadish, Tova Milo
摘要
Generalizing from detailed data to statements in a broader context is often critical for users to make sense of large data sets. Correspondingly, poorly constructed generalizations might convey misleading information even if the statements are technically supported by the data. For example, a cherry-picked level of aggregation could obscure substantial sub-groups that oppose the generalization. We present a framework for detecting and explaining cherry-picked generalizations by refining aggregate queries. We present a scoring method to indicate the appropriateness of the generalizations. We design efficient algorithms for score computation. For providing a better understanding of the resulting score, we also formulate practical explanation tasks to disclose significant counterexamples and provide better alternatives to the statement. We conduct experiments using real-world data sets and examples to show the effectiveness of our proposed evaluation metric and the efficiency of our algorithmic framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Summarized Causal Explanations For Aggregate ViewsBrit Youngmann, Michael J. Cafarella, Amir Gilad, Sudeepa RoySIGMOD 2024 · 被引用 12 次
- Visualization Guardrails: Designing Interventions Against Cherry-Picking in Interactive Data ExplorersMaxim Lisnic, Zach Cutler, Marina Kogan, Alexander LexCHI 2025 · 被引用 9 次
- On Explaining Confounding BiasBrit Youngmann, Michael J. Cafarella, Yuval Moskovitch, Babak SalimiICDE 2023 · 被引用 7 次
- Does This Have a Particular Meaning? Interactive Pattern Explanation for Network VisualizationsXinhuan Shu, Alexis Pister, Junxiu Tang, Fanny Chevalier 等IEEE VIS 2024 · 被引用 6 次
- Finding Convincing Views to Endorse a ClaimShunit Agmon, Amir Gilad, Brit Youngmann, Shahar Zoarets 等VLDB 2025 · 被引用 5 次
它引用的顶会 Paper2
相关 Paper
- Reptile: Aggregation-level Explanations for Hierarchical DataZezhou Huang, Eugene WuSIGMOD 2022 · 被引用 5 次
- DPXPlain: Privately Explaining Aggregate Query AnswersYuchao Tao, Amir Gilad, Ashwin Machanavajjhala, Sudeepa RoyVLDB 2023 · 被引用 15 次
- SDEcho: Efficient Explanation of Aggregated Sequence DifferenceFei Ye, Zikang Liu, Xi Zhang, Yinan Jing 等VLDB 2025
- Query-Guided Analysis and Mitigation of Data Verification ErrorsRan Schreiber, Yael AmsterdamerICDE 2026
- Efficient Exploration of Interesting Aggregates in RDF GraphsYanlei Diao, Pawel Guzewicz, Ioana Manolescu, Mirjana MazuranSIGMOD 2021 · 被引用 5 次
