Aws Data Engineer
CurrentAs an AWS Data Engineer at Amazon, I leverage AWS Glue to create scalable ETL processes that ensure efficient data integration across various systems. Utilizing the Hadoop ecosystem, including HDFS and MapReduce, I manage both structured and unstructured data for comprehensive analysis. I optimize data workflows on Databricks with Spark for high-speed processing and use Talend for seamless data integration while maintaining quality. My work involves engineering data pipelines on Impala and Hive for quick SQL queries, implementing workflows with Apache Airflow, and automating deployment processes using Jenkins for CI/CD. I orchestrate data flows with Nifi, develop data processing applications in Scala, and utilize NoSQL databases like MongoDB and Cassandra for flexible data modeling. Additionally, I integrate Hibernate and Spring frameworks, manage workflows with Oozie, and utilize various AWS services for resource optimization. My contributions have led to a successful migration of data processes to AWS Redshift, resulting in a 15% reduction in infrastructure costs.