The influence of dimensions on the complexity of computing decision trees

Kobourov, Stephen; Löffler, Maarten; Montecchiani, Fabrizio; Pilipczuk, Marcin; Rutter, Ignaz; Seidel, Raimund; Sorge, Manuel; Wulms, Jules

doi:https://doi.org/10.1016/j.artint.2025.104322

The influence of dimensions on the complexity of computing decision trees

Files

Southern_Economic_Journal_-_2024_-_Celis_Galvez_-_Discr... (7.74 MB)

Publication date

2025-06

Authors

Kobourov, Stephen

Löffler, Maarten

Montecchiani, Fabrizio

Pilipczuk, Marcin

Rutter, Ignaz

Seidel, Raimund

Sorge, Manuel

Wulms, Jules

DOI

https://doi.org/10.1016/j.artint.2025.104322

Document Type

Article

Metadata

Show full item record

Collections

Utrecht University Repository

License

cc_by_nc_nd

Abstract

A decision tree recursively splits a feature space Rd and then assigns class labels based on the resulting partition. Decision trees have been part of the basic machine-learning toolkit for decades. A large body of work considers heuristic algorithms that compute a decision tree from training data, usually aiming to minimize in particular the size of the resulting tree. In contrast, little is known about the complexity of the underlying computational problem of computing a minimum-size tree for the given training data. We study this problem with respect to the number d of dimensions of the feature space Rd, which contains n training examples. We show that it can be solved in O(n2d+1) time, but under reasonable complexity-theoretic assumptions it is not possible to achieve f(d)⋅no(d/log⁡d) running time. The problem is solvable in (dR)O(dR)⋅n1+o(1) time if there are exactly two classes and R is an upper bound on the number of tree leaves labeled with the first class.

Keywords

Decision trees, Machine learning, Parameterized complexity, Language and Linguistics, Linguistics and Language, Artificial Intelligence

Citation

Kobourov, S, Löffler, M, Montecchiani, F, Pilipczuk, M, Rutter, I, Seidel, R, Sorge, M & Wulms, J 2025, 'The influence of dimensions on the complexity of computing decision trees', Artificial Intelligence, vol. 343, 104322. https://doi.org/10.1016/j.artint.2025.104322

URI

https://dspace.library.uu.nl/handle/1874/471229

The influence of dimensions on the complexity of computing decision trees

Files

Publication date

Authors

Editors

Advisors

Supervisors

DOI

Document Type

Metadata

Collections

License

Abstract

Keywords

Citation

URI