Senior Data Engineer
CurrentI have a strong background in creating and refining ETL procedures, and I've worked with AWS Glue, S3, and Redshift extensively to move campaign data from a variety of sources with ease. Creating data pipelines, obtaining data from AWS S3, converting it, and importing it into AWS Redshift are all areas of my experience. Furthermore, I've skillfully integrated clickstream data from AWS Kinesis streams, using AWS S3 to store the combined outcomes before completing the load to AWS Redshift.My ability to leverage technologies like Informatica Power Center for ETL operations has allowed me to load, extract, and transform data from flat files and Oracle among other sources into Netezza Data Warehouse. Using PySpark, I have created extensive frameworks for data intake. To preserve data integrity in HIVE tables, I have carried out data aggregation, de-duplication, and cleansing. Additionally, my expertise in implementing Hadoop systems and managing NoSQL databases has been critical in adapting to changing business needs and data architectures, all the while guaranteeing effective data storage and retrieval using Kafka log compaction and retention guidelines.I have a wealth of experience integrating strong security features like role-based access control and data encryption with cutting-edge data management tools like Time Travel, Zero-Copy Cloning, and Data Sharing within Snowflake. I have experience with Power BI and Tableau for data visualization and analysis. I have planned server capacity, implemented row-level security, and made sure that organizational standards were followed. Throughout my career, I have continuously enhanced code quality and consistency by automating repetitive tasks with Git hooks and scripts, reflecting my commitment to maintaining high standards in all projects.