HUM AI SCI NA Naman Vipul Chheda GRANDE: Teaching Decision Trees to Learn Like Neural Networks How gradient descent can train entire decision tree ensembles end-to-end — and beat XGBoost.