Data Engineer
CurrentI specialize in designing, building, and optimizing scalable data pipelines and infrastructure. With extensive experience in Python and SQL, I’m adept at data manipulation, scripting, and managing complex ETL processes. My expertise spans across big data technologies like Apache Hadoop, Spark, Hive, and Kafka, as well as cloud platforms including AWS and Microsoft Azure.I have hands-on experience with data warehousing solutions such as Amazon Redshift, Google Big Query, and Snowflake, and a strong understanding of dimensional modeling for efficient data warehouse design. My technical skill set also includes working with RDBMS (MySQL, PostgreSQL, Oracle) and NoSQL databases (MongoDB, Cassandra, HBase).I’m proficient in automating data workflows using ETL tools like Apache NiFi, Talend, Informatica, and AWS Glue, and in infrastructure automation with AWS CloudFormation and Terraform. My experience extends to orchestrating data pipelines with Apache Airflow, Luigi, and Prefect, ensuring seamless data flow and processing.I’m committed to data quality and governance, familiar with tools like Great Expectations, and adhere to best practices in data security. My background also includes creating impactful data visualizations and dashboards using Tableau and Power BI, as well as deploying machine learning models within data pipelines.I thrive in collaborative environments, combining excellent problem-solving abilities with strong communication skills to drive data-driven decision-making and innovation.