Data Engineer
Current- Automated financial reporting & data quality (Python, PySpark, AWS): Streamlined processes (65% faster) with AWS EMR for the spark & improved data quality (80%).- Custom metric extraction & real-time reporting: Developed scripts for client insights, and applied risk management scripts, achieving a 60% improvement in data accuracy and integrity.- Big data & streaming: Leveraged Apache Spark and Kafka Streams to process high-volume, real-time data streams for various applications, enabling faster insights and decision-making.- Process optimization & collaboration: Implemented indexing (50% performance improvement), collaborated with stakeholders, & ensured process accuracy.