Hugging Face Trending Papers

Learned Look-Ahead Splitting Rule for CART

arXiv Machine Learning
Sep 16

Learned Look-Ahead Splitting Rule for CART

The paper introduces a look‑ahead splitting rule for Classification and Regression Trees (CART) that evaluates candidate splits by the error reduction achieved after growing a conventional CART subtree beneath each split. To keep the method computationally feasible, a smart look‑ahead algorithm is proposed that learns downstream split values from node‑level features. Experiments on simulated data and two real datasets show that both full and smart look‑ahead methods outperform the standard greedy splitting strategy, especially in hierarchical or interaction‑driven scenarios.

By Andrew Gao, Tianlin Liu, Ruichen Han, Lu Tian
Hugging Face Trending Papers
Aug 20

DICS: Data-Informed Centroid Splitting for Decision Tree Classifiers

The paper introduces DICS, a clustering-based framework that uses data-informed priors to construct a compact set of candidate splits for decision tree classifiers. By incorporating class-aware structure, DICS reduces the split search space, preserving predictive performance while cutting training time. The authors provide theoretical analysis and experimental results showing comparable accuracy to exhaustive search across synthetic and benchmark datasets.

arXiv AI
3d ago

A Moving-Horizon Approximate Branch-and-Reduce Method for Deep Classification Trees

The paper introduces a moving-horizon approximate branch‑and‑reduce method for training deep classification trees on large datasets with continuous features. It combines a hierarchical root‑subtree optimization framework, branch‑and‑reduce at the root, greedy heuristics for subtrees, and a low‑cost moving‑horizon refinement to improve accuracy. Experiments show the approach surpasses heuristic baselines in test accuracy while scaling better in dataset size and tree depth than existing global optimal solvers.

By Chenxuanyin Zou, Jiayang Ren, Qiangqiang Mao, Jing Liu, Marcus Lai, Yankai Cao
arXiv Machine Learning
Jul 3

Conditional Inference Trees and Forests for Feature Selection

arXiv:2607. 01417v1 Announce Type: new Abstract: Conditional inference trees (CIT) and conditional inference forests (CIF) reduce split-selection bias by testing features before choosing split thresholds, but repeated permutation tests and threshold searches can make these methods computationally expensive.

By Robert Milletich, Justin Downes, Steve Goley, Newel Hirst
Hugging Face Trending Papers
Jun 11

Simultaneous Latent Budget Trees for Stratified Classification

In the era of Explainable Artificial Intelligence, there is a renewed focus on single trees for their ease of interpretation. This paper introduces Simultaneous Latent Budget Trees, a probabilistic machine learning framework for classification trees in the presence of a stratification factor such as a temporal, spatial, or demographic variable, acting as a control variable or potential confounder.