Sample Complexity of Forecast Aggregation
Tao Lin, Yiling Chen
Abstract
We consider a Bayesian forecast aggregation model where n experts, after observing private signals about an unknown binary event, report their posterior beliefs about the event to a principal, who then aggregates the reports into a single prediction for the event. The signals of the experts and the outcome of the event follow a joint distribution that is unknown to the principal, but the principal has access to i.i.d. "samples" from the distribution, where each sample is a tuple of the experts' reports (not signals) and the realization of the event. Using these samples, the principal aims to find an ε-approximately optimal aggregator, where optimality is measured in terms of the expected squared distance between the aggregated prediction and the realization of the event. We show that the sample complexity of this problem is at least Ω(m n-2 /ε) for arbitrary discrete distributions, where m is the size of each expert's signal space. This sample complexity grows exponentially in the number of experts n. But, if the experts' signals are independent conditioned on the realization of the event, then the sample complexity is significantly reduced, to Õ(1/ε 2 ), which does not depend on n. Our results can be generalized to non-binary events. The proof of our results uses a reduction from the distribution learning problem and reveals the fact that forecast aggregation is almost as difficult as distribution learning. * A short version of this paper is accepted by NeurIPS 2023 (spotlight). We would like to thank Yannai Gonczarowski,
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7f6d04ec-cffc-4c05-a476-7338848fb087Cited by top-tier papers1
Ask how each one uses itBuilds on2
Related papers
- Robust Decision Aggregation with Second-order InformationYuqi Pan, Zhaohua Chen, Yuqing KongWWW 2024 · 9 citations
- Robust Aggregation with Adversarial ExpertsYongkang Guo, Yuqing KongWWW 2025 · 2 citations
- Wisdom of the Crowd Voting: Truthful Aggregation of Voter Information and PreferencesGrant Schoenebeck, Biaoshuai TaoNeurIPS 2021 · 22 citations
- Adaptive Selective Sampling for Online Prediction with ExpertsRui M. Castro, Fredrik Hellström, Tim van ErvenNeurIPS 2023 · 4 citations
- Statistically Near-Optimal Hypothesis SelectionOlivier Bousquet, Mark Braverman, Gillat Kol, Klim Efremenko et al.FOCS 2021
