Azure Cloud Data Engineer
CurrentExtract Transform and Load data from Sources Systems to Azure Data Storage services using a combination of Azure Data Factory, T-SQL, Spark SQL.Data Ingestion to one or more Azure Services - (Azure Data Lake, Azure Storage, Azure SQL, Azure DW) and processing the data in In Azure Databricks.Created Pipelines in ADF using Linked Services/Datasets/Pipeline/ to Extract, Transform and load data from different sources like Azure SQL, Blob storage, Azure SQL Data warehouse.Developed Spark applications using Pyspark and Spark-SQL for data extraction, transformation and aggregation from multiple file formats for analyzing & transforming the data to uncover insights into the customer usage patterns.Created Spark Structured Streaming Applications to consume data from EventHub and load into Delta tables after transformations.Built Spark Batch application to publish messages to Service Bus for SAP application integration with Mule Soft.Developed JSON Scripts for deploying the Pipeline in Azure Data Factory (ADF) that process the data using the Sql Activity.Hands-on experience on developing SQL Scripts for automation purpose.Created Build and Release for multiple projects (modules) in production environment using Azure DevOps.Responsible for documenting the process and cleanup of unwanted data.Hands on experience in working on Spark SQL queries, Data frames, and import data from Data sources, perform transformations; perform read/write operations, save the results to output directory into HDFS.Worked in an Agile development environment in sprint cycles of two weeks by dividing and organizing tasks.Created Databricks Delta tables to store various data formats coming from different applications.Experience in working with Restful APIs.Responsible for modifying the code, debugging, and testing the code before deploying on the production cluster.