Data Engineer
Current• Data modeling • Design and monitoring of ETL/ELT pipelines• Data wrangling using Pandas, PySpark • PySpark Performance optimization and tuning of jobs to avoid OOM and data skewness issues• Scheduling/workflow orchestration• Querying from Data Warehouse / Database using SQL