Senior Data Engineer
Current- Responsible for developing & implementing end to end solutions using Hadoop, Spark,Flume, Hive, Pig, Sqoop, Cassandra, HBase, MongoDB, spark, zookeeper, AWS.- Installed/Configured/Maintained Apache Hadoop clusters for application development andHadoop tools like Hive, Pig, HBase, Zookeeper, and Sqoop.- Managed 350+ Nodes CDH 5.2 cluster with 4 petabytes of data using Cloudera Manager andLinux RedHat 6.5.- Developed data pipeline using Flume, Sqoop, Pig, and Java map-reduce to ingest customerbehavioral data and financial histories into HDFS for analysis.- Involved in collecting and aggregating large amounts of log data using Apache Flume andstaging data in HDFS for further analysis.