Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

Approximation-Free Differentiable Oblique Decision Trees

Дата публикации: 17-08-2026 20:26:00


Decision Trees (DTs) are widely used in safety-critical domains such as medical diagnosis, valued for their interpretability and effectiveness on tabular data. However, training accurate oblique DTs is challenging due to complex optimization landscapes and overfitting risks, particularly in regression. Recent advances have introduced differentiable formulations that enable gradient-based training and joint optimization of decision boundaries and leaf regressors. Yet, existing approaches typically rely on approximations, either through probabilistic softening of boundaries (soft DTs) or quantized gradients such as the Straight-Through Estimator (STE). To overcome these limitations, we propose DTSemNet, a novel, semantically equivalent, and invertible representation of hard oblique DTs as neural networks. DTSemNet enables end-to-end training with standard gradient descent, eliminating the need for approximations in both classification and regression. While classification aligns naturally with this formulation, regression remains challenging due to the joint optimization of internal nodes and leaf regressors. To address this, we analyze the limitations of STE and introduce an annealed Top-$k$ method that provides accurate gradient signals without approximation. Extensive experiments on classification and regression benchmarks show that DTSemNet-trained oblique DTs outperform state-of-the-art differentiable DTs. Furthermore, we demonstrate that DTSemNet can serve as programmatic DT policies in reinforcement learning environments, thereby broadening their applicability.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1 End-to-End Deep Learning for Predicting Metric Space-Valued Outputs 010.6617-08-2026
2 Abstract Gradient Training: A Unified Certification Framework for Data Poisoning, Unlearning, and Differential Privacy 07.1717-08-2026
3 High-Dimensional Analysis of Gradient Flow for Extensive-Width Quadratic Neural Networks 08.717-08-2026
4 Minimax Optimal Convergence of Gradient Descent in Logistic Regression via Large and Adaptive Stepsizes 07.5217-08-2026
5 Near-optimal Delta-convex Estimation of Lipschitz Functions 09.7117-08-2026
6 Neural Exploitation and Exploration of Contextual Bandits 06.3417-08-2026
7 The Sample Complexity of Parameter-Free Stochastic Convex Optimization 05.717-08-2026
8 Limiting Over-Smoothing and Over-Squashing of Graph Message Passing by Deep Scattering Transforms 010.8717-08-2026
9 Graph-based Clustering Revisited: A Relaxation of Kernel k-Means Perspective 010.9417-08-2026
10 Doubly Debiased Robust Subsampling for Transfer Learning 08.0617-08-2026

Классификация: Пресс-релизы. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 9.6. Источник: jmlr.org.