FedUP: Querying Large-Scale Federations of SPARQL Endpoints
Julien Aimonier-Davat, Brice Nédelec, Minh Hoang Dang, Pascal Molli, Hala Skaf-Molli
Abstract
Processing SPARQL queries over large federations of SPARQL endpoints is crucial for keeping the Semantic Web decentralized. Despite the existence of hundreds of SPARQL endpoints, current federation engines only scale to dozens. One major issue comes from the current definition of the source selection problem, i.e., finding the minimal set of SPARQL endpoints to contact per triple pattern. Even if such a source selection is minimal, only a few combinations of sources may return results. Consequently, most of the query processing time is wasted evaluating combinations that return no results. In this paper, we introduce the concept of Result-Aware query plans. This concept ensures that every subquery of the query plan effectively contributes to the result of the query. To compute a Result-Aware query plan, we propose FedUP, a new federation engine able to produce Result-Aware query plans by tracking the provenance of query results. However, getting query results requires computing source selection, and computing source selection requires query results. To break this vicious cycle, FedUP computes results and provenances on tiny quotient summaries of federations at the cost of source selection accuracy. Experimental results on federated benchmarks demonstrate that FedUP outperforms state-of-the-art federation engines by orders of magnitude in the context of large-scale federations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c3693f92-f3c0-4b02-bece-0ffc10a79babCited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- Accio: Bolt-on Query FederationXiaoying Wang, Jiannan Wang, Tianzheng Wang, Yong ZhangVLDB 2025
- NPCS: Native Provenance Computation for SPARQLZubaria Asma, Daniel Hernández, Luis Galárraga, Giorgos Flouris et al.WWW 2024 · 4 citations
- Computing How-Provenance for SPARQL Queries via Query RewritingDaniel Hernández, Luis Galárraga, Katja HoseVLDB 2021 · 42 citations
- Efficient Execution of SPARQL Queries with OPTIONAL and UNION ExpressionsYue Pang, Lei Zou, M. Tamer Özsu, Jiaqi ChenICDE 2025
- WiseKG: Balanced Access to Web Knowledge GraphsAmr Azzam, Christian Aebeloe, Gabriela Montoya, Ilkcan Keles et al.WWW 2021 · 21 citations
