Abstract
Federated learning allows training machine learning (ML) models without central consolidation of the raw data. Variants of such federated learning systems enable privacy-preserving ML, and address data ownership and/or sharing constraints. However, existing work mostly adopt data-parallel parameter-server architectures for mini-batch training, require manual construction of federated runtime plans, and largely ignore the broad variety of data preparation, ML algorithms, and model debugging. Over the last years, we extended Apache SystemDS by an additional federated runtime backend for federated linear-algebra programs, federated parameter servers, and federated data preparation. In this paper, we share the system-level compiler and runtime integration, new features such as multi-tenant federated learning, selected federated primitives, multi-key homomorphic encryption, and our monitoring infrastructure. Our demonstrator showcases how composite ML pipelines can be compiled into federated runtime plans with low overhead.
Author supplied keywords
Cite
CITATION STYLE
Baunsgaard, S., Boehm, M., Innerebner, K., Kehayov, M., Lackner, F., Ovcharenko, O., … Wrede, S. B. (2022). Federated Data Preparation, Learning, and Debugging in Apache SystemDS. In International Conference on Information and Knowledge Management, Proceedings (pp. 4813–4817). Association for Computing Machinery. https://doi.org/10.1145/3511808.3557162
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.