Lune

ICML2024顶会

Fully-Dynamic Approximate Decision Trees With Worst-Case Update Time Guarantees

Marco Bressan, Mauro Sozio

2024年份
1被引次数

摘要

We give the first algorithm that maintains an approximate decision tree over an arbitrary sequence of insertions and deletions of labeled examples, with strong guarantees on the worst-case running time per update request. For instance, we show how to maintain a decision tree where every vertex has Gini gain within an additive α\alpha of the optimum by performing O(d (log⁡n)4α3)O\Big(\frac{d\,(\log n)^4}{\alpha^3}\Big) elementary operations per update, where dd is the number of features and nn the maximum size of the active set (the net result of the update requests). We give similar bounds for the information gain and the variance gain. In fact, all these bounds are corollaries of a more general result, stated in terms of decision rules -- functions that, given a set SS of labeled examples, decide whether to split SS or predict a label. Decision rules give a unified view of greedy decision tree algorithms regardless of the example and label domains, and lead to a general notion of ϵ\epsilon-approximate decision trees that, for natural decision rules such as those used by ID3 or C4.5, implies the gain approximation guarantees above. The heart of our work provides a deterministic algorithm that, given any decision rule and any ϵ>0\epsilon>0, maintains an ϵ\epsilon-approximate tree using O ⁣(d f(n)npoly⁡hϵ)O\!\left(\frac{d\, f(n)}{n} \operatorname{poly}\frac{h}{\epsilon}\right) operations per update, where f(n)f(n) is the complexity of evaluating the rule over a set of nn examples and hh is the maximum height of the maintained tree.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper3

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖