Steering Semantic Data Processing With DocWrangler
Shreya Shankar, Bhavya Chopra, Mawil Hasan, Stephen Lee, Bjoern Hartmann, Joseph M. Hellerstein, Aditya G. Parameswaran, Eugene Wu
2025Year
3Top-tier citations
Abstract
intent is hard to communicat Prompts need to be detailed and dataspecific LLM behavior varies across document Requires fine-grained decomposition A B C Docs & Outputs condition discomfort_level symptoms
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f6652737-044a-4d58-a485-7c69b0093e63Cited by top-tier papers3
- KRAMABENCH: A Benchmark for AI Systems on Data-to-Insight Pipelines over Data LakesEugenie Lai, Gerardo Vitagliano, Ziyu Zhang, Om Chabra et al.ICLR 2026 · 37 citations
- Multi-Objective Agentic Rewrites for Unstructured Data ProcessingLindsey Linxi Wei, Shreya Shankar, Sepanta Zeighami, Yeounoh Chung et al.VLDB 2026 · 15 citations
- SEMA: A High-performance System for LLM-based Semantic Query ProcessingKangkang Qi, Dongyang Xie, Wenbo Li, Hao Zhang et al.VLDB 2026 · 5 citations
Builds on33
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan et al.NeurIPS 2023 · 4,972 citations
- Design Guidelines for Prompt Engineering Text-to-Image Generative ModelsVivian Liu, Lydia B. ChiltonCHI 2022 · 586 citations
- AI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model PromptsTongshuang Wu, Michael Terry, Carrie Jun CaiCHI 2022 · 465 citations
- The Metacognitive Demands and Opportunities of Generative AILev Tankelevitch, Viktor Kewenig, Auste Simkute, Ava Elizabeth Scott et al.CHI 2024 · 279 citations
Related papers
- Be Responsible in Your Answers! Monitoring Out-of-Domain Behaviors in Domain-Specific LLMsBoquan Li, Chenzhe Lou, Zhe Ren, Peixin Zhang et al.WWW 2026
- MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical ReasoningShuyue Stella Li, Vidhisha Balachandran, Shangbin Feng, Jonathan Ilgen et al.NeurIPS 2024 · 215 citations
- Live in the Loop: Rapid Run-time Feedback for PromptsToni Mattis, Abdullatif Ghajar, Tom Beckmann, Robert HirschfeldCHI 2026 · 2 citations
- Measuring Intent Comprehension in LLMsNadav Kunievsky, James EvansICML 2026 · 1 citation
- Ask and Retrieve Knowledge: Towards Proactive Asking with Imperfect Information in Medical Multi-turn DialoguesBolin Zhang, Shengwei Wang, Yangqin Jiang, Dianbo Sui et al.SIGIR 2025
