cs.LGNov 6, 2025

Towards Scalable Meta-Learning of near-optimal Interpretable Models via Synthetic Model Generations

Authors: Kyaw Hpone Myint, Zhe Wu, Alexandre G. R. Day, Giri Iyengar

Organizations: Capital One

Abstract

Decision trees are widely used in high-stakes fields like finance and healthcare due to their interpretability. This work introduces an efficient, scalable method for generating synthetic pre-training data to enable meta-learning of decision trees. Our approach samples near-optimal decision trees synthetically, creating large-scale, realistic datasets. Using the MetaTree transformer architecture, we demonstrate that this method achieves performance comparable to pre-training on real-world data or with computationally expensive optimal decision trees. This strategy significantly reduces computational costs, enhances data generation flexibility, and paves the way for scalable and efficient meta-learning of interpretable decision tree models.

Figures & tables

Appendix figures & tables2 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. MotherTree: Meta-learning on synthetic data improves decision tree training

    Oct 7, 2026Ziyuan Wang, Fredrik D. JohanssonDecision Tree LearningMeta-Learning

  2. Multistage Defer Trees for Hybrid Interpretability: If at First You Can't Succeed, Tree Again

    Jun 30, 2026Zakk Heile, Hayden McTavish, Margo Seltzer +1Ensemble LearningLearning to Defer

  3. From Prompts to Trees: Effective LLM-Guided Tree Generation for Few-Shot Tabular Classification

    Oct 7, 2026Yue Qiu, Zekang Du, Yiqun Diao +2Tabular ClassificationDecision Tree Learning