Current browse context:
stat.ML
Change to browse by:
References & Citations
Statistics > Machine Learning
Title: Riemann-Lebesgue Forest for Regression
(Submitted on 7 Feb 2024 (v1), last revised 10 May 2024 (this version, v3))
Abstract: We propose a novel ensemble method called Riemann-Lebesgue Forest (RLF) for regression. The core idea in RLF is to mimic the way how a measurable function can be approximated by partitioning its range into a few intervals. With this idea in mind, we develop a new tree learner named Riemann-Lebesgue Tree (RLT) which has a chance to perform Lebesgue type cutting,i.e splitting the node from response $Y$ at certain non-terminal nodes. We show that the optimal Lebesgue type cutting results in larger variance reduction in response $Y$ than ordinary CART \cite{Breiman1984ClassificationAR} cutting (an analogue of Riemann partition). Such property is beneficial to the ensemble part of RLF. We also generalize the asymptotic normality of RLF under different parameter settings. Two one-dimensional examples are provided to illustrate the flexibility of RLF. The competitive performance of RLF against original random forest \cite{Breiman2001RandomF} is demonstrated by experiments in simulation data and real world datasets.
Submission history
From: Tian Qin [view email][v1] Wed, 7 Feb 2024 03:13:11 GMT (181kb,D)
[v2] Wed, 8 May 2024 16:51:08 GMT (446kb,D)
[v3] Fri, 10 May 2024 01:06:41 GMT (439kb,D)
Link back to: arXiv, form interface, contact.