Data Engineer
Current• Develop and maintain Apache Hadoop, Hive, Spark, Clickhouse, and Trino on top of Kubernetes as infrastructure using Helm charts.• Manage Data pipelines that enables us to integrate data from different data sources (such as Postgresql, Kafka, and etc) into some destinations (such as HDFS) to build a data warehouse or a datalake system.