Sr Data Engineer
CurrentAs a data engineer, I have developed and managed ETL workflows using AWS Data Pipeline, AWS Glue, and Informatica PowerCenter, ensuring efficient data integration and processing for large datasets. My expertise includes designing scalable data pipelines, integrating diverse data sources into data warehouses, and optimizing data processing through tools like Talend and Matillion.I have extensive experience with Apache Spark, implementing PySpark jobs and developing data ingestion workflows using Sqoop and Flume. My focus on performance optimization and fault tolerance has enabled me to architect complex ETL pipelines with Apache Airflow, while also managing SQL Server instances and migrating databases to Snowflake and Redshift.In addition to technical development, I have created interactive Power BI dashboards, provided training on best practices, and implemented CI/CD workflows using GitHub Actions. My commitment to documentation and process improvement has facilitated smoother deployment and enhanced collaboration across teams