AeroLeads people directory · profile

Naveen Kumar Email & Phone Number

Azure Databricks Lead Data Engineer at Cencora
Location: Jersey City, New Jersey, United States 7 work roles 1 school
LinkedIn matched
✓ Verified August 2026 3 data sources Profile completeness 86%

Contact Signals

LinkedIn Profile matched
3 free lookups remaining · No credit card
Current company
Role
Azure Databricks Lead Data Engineer
Location
Jersey City, New Jersey, United States
Company size

Who is Naveen Kumar? Overview

A concise factual answer block for searchers comparing this professional profile.

Quick answer

Naveen Kumar is listed as Azure Databricks Lead Data Engineer at Cencora, a with 26687 employees, based in Jersey City, New Jersey, United States. AeroLeads shows a matched LinkedIn profile for Naveen Kumar.

Naveen Kumar previously worked as Azure Databricks Lead Data Engineer at Bcbs Fl and Azure Databricks Senior Data Engineer at Bcbs Fl. Naveen Kumar holds Bachelor Of Technology - Btech, Electrical, Electronics And Communications Engineering from Acharya Nagarjuna University.

Company email context

Email format at Cencora

This section adds company-level context without repeating Naveen Kumar's masked contact details.

Cencora

Review company-level records connected to Naveen Kumar before choosing the right outreach path.

Profile bio

About Naveen Kumar

Naveen Kumar is a Azure Databricks Lead Data Engineer at Cencora.

Current workplace

Naveen Kumar's current company

Company context helps verify the profile and gives searchers a useful next step.

Cencora
Cencora
Azure Databricks Lead Data Engineer
Jersey City, NJ, US
Website
Employees
26687
AeroLeads page
7 roles

Naveen Kumar work experience

A career timeline built from the work history available for this profile.

Azure Databricks Lead Data Engineer

Jersey City, Nj, Us

Azure Databricks Lead Data Engineer

Bcbs Fl

Jersey City, Nj, Us

Azure Databricks Lead Data Engineer

Analyze, design, and build modern data solutions using Azure PaaS service to support visualization of data. Understand current Production state of application and determine the impact of new implementations on existing business processes.Worked with Businesses Analyst, Technical Architects and Project Managers on a daily basis to get the project requirements and analyzed these requirements and then converted the business document to a technical requirement document that enlists all the technical architecture and components of requirement.Lead the team of 4 which includes both onsite and offsite resources in developing the new or maintaining existing pipelines.Created pipelines in Azure Data Factory using linked services to various source systems and then used databricks notebook for transformation and then loaded the data into azure blob storage.Created multiple realtime data ingestion scripts and batch ingestion scripts using PySpark to meet specific business requirements.Developed Databricks Asset Bundle yaml scripts for deploying pipelines in Azure Databricks.Performed the optimisation using broadcast variables and broadcast joins and used incremental loads for job processing.Developed SQL Scripts for data validation in databricks.Used PySpark and Spark-SQL for data extraction, transformation, and aggregation from multiple file formats to uncover customer usage patterns.Creating data models for bringing in the new data source for supporting various processes.Co-ordinated with various teams for on time release code into production environment.Environment: Azure Databricks, Azure Data Factory, Azure Blob Storage, Azure Logic Apps, Pyspark, SQL

Azure Databricks Senior Data Engineer

Bcbs Fl

Client: BCBS FLResponsibilities:Analyze, design and build modern data solutions using Azure PaaS service to support visualization of data. Understand current Production state of application and determine the impact of new implementations on existing business processes.Data Ingestion Azure storage services and processing the data in azure databricks.Created pipelines in Azure Data Factory using linked services to databricks for extraction, transformation and loading data from various azure storage sources.Created UDFs in Scala and PySpark to meet specific business requirements. Develop JSON scripts for deploying pipelines in Azure Data Factory (ADF) that process data using SQL activities.Developed SQL Scripts for data validation in databricks.Used PySpark and Spark-SQL for data extraction, transformation, and aggregation from multiple file formats to uncover customer usage patterns.Creating data model for bringing in the new data source for supporting various process.

Aug 2021 - Jun 2023

Data Engineer

Schaumburg, Il, Us

Client Description:Zurich is a global insurance company which is organized into three core business segments: General Insurance, Global Life and Farmers. Zurich North American Insurance is an commercial property-casualty insurance provider. Its solutions server diverse set of industries including automotive, construction, manufacturing, technology and numerous others.Responsibilities:Migration of ETL process into Hadoop using spark and Scala.Optimizing of existing algorithms in Hadoop using Spark Context, Spark-SQL, Data Frames and Pair RDD's.Used HiveQL for migrating the smaller legacy process into meld.Bringing streaming data using Apache Kafka and Spark StreamingBuilding the required data lake in Hadoop using Sqoop from various data sourceExperienced in developing custom input formats and data types to parse and process unstructured and semi structured input data and mapped them into key value pairs to implement business logic in Map-Reduce.Creating data model for bringing in the new data source for supporting various process.Bundling and compiling the code for creating jars using SBTUnit testing on data and improvement of performance, turned over to production

Jun 2018 - Jul 2021

Data Engineer

Us

Project 1: Cargo-jet Airways Ltd, ON, Canada - RemoteResponsibilities:Installed and configured Hadoop Mapreduce HDFS Developed multiple MapReduce jobs in java for data cleaning and preprocessing.Installed and configured Pig and also written PigLatin scripts.Involved in managing and reviewing Hadoop log files.Imported data using Sqoop to load data from MySQL to HDFS on regular basis.Developing Scripts and Batch Job to schedule various Hadoop Program.Creating Hive tables and working on them using Hive QL.Importing and exporting data into HDFS and Hive using Sqoop.Involved in creating Hive tables loading with data and writing hive queries which will run internally in map reduce way.Developed a custom FileSystem plug in for Hadoop so it can access files on Data Platform. This plugin allows MapReduce programs HBase Pig and Hive to work unmodified and access files directly.Designed and implemented Mapreduce-based large-scale parallel relation-learning systemExtracted feeds form social media sites such as Facebook Twitter using Python scripts.Setup and benchmarked HBase clusters for internal use Project 2 : Charles Schwab, TXResponsibilities:This role includes working various application team to ensure technical integrity and consistency of solutions. Creation of Queries to take in a business request for a new data extract and then code a statement that will return the desired values.Performed requirement gathering and specifications analysis with clients and implementing POCs.Processed data in HDFS by developing solutions, analyzed the data using mapreduce, Hive and produce summary results from hadoop to downstream systems.Created Flows to ingest data from various systems/sources to HDFSInvolved in creating tables and then apply HQL on these table for data validationMoved data from hive tables to mongo DB Reviewed Hadoop log files.Involved in loading and transforming large sets of data and analyze them using Hive queries and scripts.

Sep 2016 - May 2018

Big Data Developer

Mechelen, Be

Description:Telenet Group is the largest provider of cable broadband services in Belgium. Its business comprises the provision of analog and digital cable television, fixed and mobile telephone services, primarily to residential customers in Flanders and Brussels. ADT is the application that was built on Customer Service Framework and it has capability to manage B2C and B2B interactions. Responsibilities: This application consolidates relevant customer information from Telenet systems, interaction data, and service requests into a composite view.They dynamically display customer data based on the customer context and current situation.Developed multiple MapReduce jobs in java for data cleaning and preprocessing involved in importing data into HDFS and Hive.Defined job flows.Involved in loading and transforming large sets of structured and unstructured data.Loaded data from UNIX system to HDFS

Sep 2014 - Aug 2016
1 education record

Naveen Kumar education

  • Acharya Nagarjuna University
    Acharya Nagarjuna University
    Electronics And Communications Engineering
FAQ

Frequently asked questions about Naveen Kumar

Quick answers generated from the profile data available on this page.

What company does Naveen Kumar work for?

Naveen Kumar works for Cencora.

What is Naveen Kumar's role at Cencora?

Naveen Kumar is listed as Azure Databricks Lead Data Engineer at Cencora.

Where is Naveen Kumar based?

Naveen Kumar is based in Jersey City, New Jersey, United States while working with Cencora.

What companies has Naveen Kumar worked for?

Naveen Kumar has worked for Cencora, Bcbs Fl, Zurich North America, Multiple Companies, and Telenet.

How can I contact Naveen Kumar?

You can use AeroLeads to view verified contact signals for Naveen Kumar at Cencora, including work email, phone, and LinkedIn data when available.

What schools did Naveen Kumar attend?

Naveen Kumar holds Bachelor Of Technology - Btech, Electrical, Electronics And Communications Engineering from Acharya Nagarjuna University.

Find 750M verified contacts

Search by job title, company, industry, location, and seniority. Export verified B2B contact data when you need it.