Abstract
The optimization of ETL (Extract, Transform, Load) pipelines using Apache Spark and Snowflake. Apache Spark is a powerful open-source distributed data processing platform, while Snowflake is a cloud-native data warehousing solution. It discusses the challenges and solutions in tuning Spark configurations using machine learning techniques and optimizing Snowflake's architecture for cost efficiency and performance. Experimental results demonstrate significant performance gains and cost savings through these optimizations.
Cite
CITATION STYLE
Mantri, A. (2023). Advanced ML (Machine Learning) Techniques for Optimizing ETL Workflows with Apache Spark and Snowflake. Journal of Artificial Intelligence & Cloud Computing, 2(3), 1. https://doi.org/10.47363/jaicc/2023(2)339
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.