Big Data Engineer
Current• Hands in experience in working with Continuous Integration and Deployment (CI/CD) using Jenkins, Docker.• Queried multiple databases like Snowflake, UDB and MySQL for data processing.• Analyzed and processed complex data sets using advanced querying using Presto, HIVE and Teradata, visualization using Tableau and analytics tools such as Python and SAS.• Developed ETL pipelines in and out data warehouse using combination of Python and Snowflake SnowSQL. Writing SQL quires against Snowflake. • Exposure to Full Lifecycle (SDLC) of Data Warehouse projects including Dimensional Data Modeling.• Worked with building data warehouse structures, and creating facts, dimensions, aggregate tables, by dimensional modeling, Star and Snowflake schemas.• Extract Transform and Load data from Sources Systems to Azure Data Storage services using a combination of Azure Data Factory, Spark SQL, and U-SQL Azure Data Lake Analytics. Data Ingestion to one or more Azure Services - (Azure Data Lake, Azure Storage, Azure SQL, Azure DW) and processing the data in In Azure Databricks.• Developed JSON Scripts for deploying the Pipeline in Azure Data Factory (ADF) that process the data using the Sql Activity.• Created Sessions and extracted data from various sources, transformed data according to the requirement and loading into data warehouse.• Used various transformations like Filter, Expression, Sequence Generator, Update Strategy, Joiner, Router and Aggregator to create robust mappings in the Informatica Power Center Designer.• Used Informatica Power Center for (ETL) extraction, transformation and loading data from heterogeneous source systems into target database.• Used ER Studio for Creating/Updating Data Models.• Creating Databricks notebooks using SQL, Python and automated notebooks using jobs.• Created mappings using Designer and extracted data from various sources, transformed data according to the requirement.