Data Engineer
CurrentDesigned and implemented resilient data pipelines for finance portfolio management using PySpark, Apache Airflow,Azure Data Factory, Azure Functions, and Apache Spark for seamless ETL processes with large financial datasets.Managed data storage management within the financial sector, optimizing performance and maintaining data integrityusing Azure Data Lake and PostgreSQL, tailored specifically to portfolio management requirements.Worked on real-time data processing initiatives, employing PySpark Streaming and Kafka Streams to analyze highvolume financial data streams, enabling timely risk mitigation strategies within the portfolio management framework.Utilized advanced data analysis and machine learning techniques with Python libraries such as Pandas and NumPy toextract actionable insights from financial data, supporting strategic decision-making in portfolio management.Developed interactive dashboards using Power BI and crafted custom visualizations with Matplotlib to effectivelycommunicate critical findings and trends to stakeholders in the finance risk management domain.