Azure Data Engineer
CurrentDesigned and Developed Azure Data Factory Pipelines to load data from various source systems into Azure data lake storage and to process it into Azure Databricks. Automated the data pipeline with Azure Data Factory (ADF), Increasing operational efficiency by 20%. Used Apache Spark to creating Spark SQL and PySpark scripts to clean and transform data into the cleaning layer. Have strong knowledge of data warehousing concepts such as CDC and SCD types 1 and 2. Worked with Spark SQL queries, data frames, imported data from data sources, performed transformations, executed read/write operations, and saved results to the output directory. Created and maintained documentation on pipeline architecture and processes for easy troubleshooting and future scalability. Collaborated with business analysts and stakeholders to understand and translate requirements into technical solutions.