Sai Chandra
AeroLeads people directory · profile

Sai Chandra Email & Phone Number

Senior Data Engineer at Calpine
Location: United States 7 work roles
LinkedIn matched
✓ Verified July 2026 2 data sources Profile completeness 71%

Contact Signals

LinkedIn Profile matched
3 free lookups remaining · No credit card
Current company
Role
Senior Data Engineer
Location
United States
Company size

Who is Sai Chandra? Overview

A concise factual answer block for searchers comparing this professional profile.

Quick answer

Sai Chandra is listed as Senior Data Engineer at Calpine, a with 2309 employees, based in United States. AeroLeads shows a matched LinkedIn profile for Sai Chandra.

Sai Chandra previously worked as Data Engineer at Prudential Financial and Data Engineer at Ditech Group.

Company email context

Email format at Calpine

This section adds company-level context without repeating Sai Chandra's masked contact details.

Calpine

Review company-level records connected to Sai Chandra before choosing the right outreach path.

Profile bio

About Sai Chandra

• Over 9+ years of IT experience in a variety of industries working on Big Data technology using technologies such as Cloudera and Hortonworks distributions. Hadoop working environment includes Hadoop, Spark, MapReduce, Kafka, Hive, Ambari, Sqoop, HBase, and Impala. Can work parallelly in both GCP and Azure Clouds coherently. • Hands on experience with different programming languages such as Python, SAS. • Able to use Sqoop to migrate data between RDBMS, NoSQL databases and HDFS. • Experience in Extraction, Transformation and Loading (ETL) data from various sources into Data Warehouses, as well as data processing like collecting, aggregating and moving data from various sources using Apache Flume, Kafka, PowerBI and Microsoft SSIS. • Hands-on experience with Hadoop architecture and various components such as Hadoop File System HDFS, Job Tracker, GCP, Task Tracker, Name Node, Data Node and Hadoop MapReduce programming. • Experience in handling python and spark context when writing Pyspark programs for ETL. • Seasoned practice in Machine Learning algorithms and Predictive Modeling such as Linear Regression, Logistic Regression, Naïve Bayes, Decision Tree, Random Forest, KNN, Neural Networks, and K-means Clustering. • Ample knowledge of data architecture including data ingestion pipeline design, Hadoop/Spark architecture, data modeling, data mining, machine learning and advanced data processing. • Implementations of generalized solution model using AWS SageMaker • Hands-on experience with Amazon EC2, Amazon S3, Amazon RDS, VPC, IAM, Amazon Elastic Load Balancing, Auto Scaling, CloudWatch, SNS, SES, SQS, Lambda, EMR and other services of the AWS family. • Used IDEs like Eclipse, IntelliJ IDE, PyCharm IDE, Notepad ++, and Visual Studio for development. • Very keen in knowing newer techno stack that Google Cloud platform (GCP) adds. • Fluent programming experience with Scala, Python, SQL, T-SQL, R. • Hands-on experience in developing and deploying enterprise-based applications using major Hadoop ecosystem components like MapReduce, YARN, Hive, HBase, Flume, Sqoop, Spark MLlib, Spark GraphX, Spark SQL, Kafka. • Worked on Dimensional Data modelling in Star and Snowflake schemas and Slowly Changing Dimensions (SCD). • Very keen in knowing newer techno stack that Google Cloud platform (GCP) adds. • Experience working with NoSQL databases like Cassandra and HBase and developed real-time read/write access to very large datasets via HBase.

Current workplace

Sai Chandra's current company

Company context helps verify the profile and gives searchers a useful next step.

Calpine
Calpine
Senior Data Engineer
United States
Website
Employees
2309
AeroLeads page
7 roles

Sai Chandra work experience

A career timeline built from the work history available for this profile.

Senior Data Engineer

United States

Data Engineer

• Experience in building and architecting multiple Data pipelines, end to end ETL and ELT process for Data ingestion and transformation in GCP and coordinate task among the team. • The roles include creating data pipelines from application databases to Data Lake and warehouse and create stream and batch data processing pipelines. • Build data pipelines in airflow in GCP for ETL related jobs using different airflow operators. • Created python scripts to ingest data from on-premise to GCS and built data pipelines using Apache Beam and Data Flow for data transformation from GCS to Big query • Maintaining more than 16 PB, 300 nodes Cloudera's distribution Hadoop production and dev clusters. Perform daily health checks, work on alerts, and other related tasks. • Created Hive databases, tables and queries for data analytics as and when needed, integrate them with data processing jobs, and drop them as a part of cleanup. Also working on maintaining HBase database used by other applications. • Leveraged Google Cloud Platform Services to process and manage the data from streaming and file-based sources. • Designed and developed in house database visualization tool built on Oracle Application Express platform that visualize the real-time database health information as well as shows management reports. • Experience in moving data between GCP and Azure using Azure Data Factory. • Experienced as AWS cloud engineer to create and manage EC2 instances, S3 buckets, and configure RDS instances. Monitoring the performance, spinning off the instances regularly, and help developers to get access to the cloud instances.

Senior Data Engineer

Dania, Florida, United States

• Worked on AWS Data pipeline to configure data loads from S3 to into Redshift. • Using AWS Redshift, I Extracted, transformed and loaded data from various heterogeneous data sources and destinations • Created Tables, Stored Procedures, and extracted data using T-SQL for business users whenever required. • Performs data analysis and design, and creates and maintains large, complex logical and physical data models, and metadata repositories using ERWIN and MB MDR • I have written shell script to trigger data Stage jobs. • Assist service developers in finding relevant content in the existing reference models. • Like Access, Excel, CSV, Oracle, flat files using connectors, tasks and transformations provided by AWS Data Pipeline. • Utilized Spark SQL API in PySpark to extract and load data and perform SQL queries. • Worked on developing Pyspark script to encrypting the raw data by using Hashing algorithms concepts on client specified columns. • Responsible for Design, Development, and testing of the database and Developed Stored Procedures, Views, and Triggers • Developed Python-based API (RESTful Web Service) to track revenue and perform revenue analysis. • Compiling and validating data from all departments and Presenting to Director Operation. • KPI calculator Sheet and maintain that sheet within SharePoint.

Jul 2019 - Sep 2021

Data Engineer

New York, United States

• Designed and Implemented Sharding and Indexing Strategies for MongoDB servers. • Optimizing pig scripts, user interface analysis, performance tuning and analysis. • Used Hive to analyze the partitioned and bucketed data and compute various metrics for reporting on the dashboard. • Implemented the import and export of data using XML and SSIS. • Involved in Planning, Defining and Designing data base using ER Studio on business requirement and provided documentation. • Responsible for migration of application running on premise onto Azure cloud. • Used SSIS to build automated multi-dimensional cubes. • Wrote indexing and data distribution strategies optimized for sub-second query response • Developed a statistical model using artificial neural networks for ranking the students to better assist the admission process. • Wrote scripts and indexing strategy for a migration to Confidential Redshift from SQL Server and MySQL databases. • Implement software enhancements to port legacy software systems to Spark and Hadoop ecosystems on Azure Cloud.

Feb 2017 - Jun 2019

Data Engineer

Hyderabad, Telangana, India

• Deployed Lambda and other dependencies into AWS to automate EMR Spin for Data Lake jobs • Scheduled spark applications/Steps in AWS EMR cluster. • Extensively used event-driven and scheduled AWS Lambda functions to trigger various AWS resources. • Implemented advanced procedures like text analytics and processing using teh in-memory computing capabilities like Apache Spark written in Scala. • Developed spark applications for performing large scale transformations and denormalization of relational datasets. • Developed and executed a migration strategy to move Data Warehouse from SAP to AWS Redshift. • Loaded data into teh cluster from dynamically generated files using Flume and from relational database management systems using Sqoop. • Used Spark Streaming to divide streaming data into batches as an input to spark engine for batch processing. • Worked on analyzing Hadoop cluster and different Big Data analytic tools including Pig, hive, HBase, Spark and Sqoop. • Exported data from HDFS to RDBMS via Sqoop for Business Intelligence, visualization, and user report generation. • Loading teh data from multiple Data sources like (SQL, DB2, and Oracle) into HDFS using Sqoop and load into Hive tables. • Performed Real time event processing of data from multiple servers in teh organization using Apache Storm by integrating wif apache Kafka. • Performed Impact Analysis of teh changes done to teh existing mappings and provided teh feedback • Involved in complete Implementation lifecycle, specialized in writing custom MapReduce, and Hive • Extensively used Hive/HQL or Hive queries to query or search for a string in Hive tables in HDFS

Apr 2015 - Nov 2016

Data Analyst

Hyderabad, Telangana, India

Involved in implementation of the project went through several phases namely: data set analysis, preprocessing data set, user-generated data extraction, and modeling.Participated in Data Acquisition with the Data Engineer team to extract historical and real-time data by using Sqoop, Pig, Flume, Hive, MapReduce, Land HDES.Wrote user-defined functions (UDFs) in Hive to manipulate strings, dates, and other data.Performed Data Cleaning, features scaling, features engineering using pandas, and NumPy packages in python .Process Improvement: Analyzed error data of recurrent programs using Python and devised a new process to reduce the turnaround time of the problem's solutions by 60%Worked on production data fixes by creating and testing SQL scripts.Deep dived into complex data sets to analyze trends using Linear Regression, Logistic Regression, Decision Trees.Prepared reports using SQL and Excel to track the performance of websites and apps.Performed Data Collection, Data Cleaning, Data Visualization, and Feature Engineering using Python libraries such as Pandas, Numpy, MatPlotLib, and seaborn.

Aug 2013 - Mar 2015
Team & coworkers

Colleagues at Calpine

Other employees you can reach at calpine.com. View company contacts for 2309 employees →

FAQ

Frequently asked questions about Sai Chandra

Quick answers generated from the profile data available on this page.

What company does Sai Chandra work for?

Sai Chandra works for Calpine.

What is Sai Chandra's role at Calpine?

Sai Chandra is listed as Senior Data Engineer at Calpine.

Where is Sai Chandra based?

Sai Chandra is based in United States while working with Calpine.

What companies has Sai Chandra worked for?

Sai Chandra has worked for Calpine, Prudential Financial, Ditech Group, Chewy, and Macy'S.

Who are Sai Chandra's colleagues at Calpine?

Sai Chandra's colleagues at Calpine include Eulete Patrick, Katherine Potter, Adam Miller, Ben Noel, and Pherone Oates.

How can I contact Sai Chandra?

You can use AeroLeads to view verified contact signals for Sai Chandra at Calpine, including work email, phone, and LinkedIn data when available.

Find 750M verified contacts

Search by job title, company, industry, location, and seniority. Export verified B2B contact data when you need it.

People with similar names

Check these profiles if this is not the Sai Chandra you were looking for.

View similar profiles