Tree-based models are among the most efficient machine learning techniques for data mining nowadays due to their accuracy, interpretability, and simplicity. The recent orthogonal needs for more data and privacy protection call for collaborative privacy-preserving solutions. In this work, we survey the literature on distributed and privacy-preserving training of tree-based models and we systematize its knowledge based on four axes: the learning algorithm, the collaborative model, the protection mechanism, and the threat model. We use this to identify the strengths and limitations of these works and provide for the first time a framework analyzing the information leakage occurring in distributed tree-based model learning.
CITATION STYLE
Chatel, S., Pyrgelis, A., Troncoso-Pastoriza, J. R., & Hubaux, J.-P. (2021). SoK: Privacy-Preserving Collaborative Tree-based Model Learning. Proceedings on Privacy Enhancing Technologies, 2021(3), 182–203. https://doi.org/10.2478/popets-2021-0043
Mendeley helps you to discover research relevant for your work.