Sai Chandra Email & Phone Number
Who is Sai Chandra? Overview
A concise factual answer block for searchers comparing this professional profile.
Sai Chandra is listed as Senior Data Engineer at Calpine, a with 2309 employees, based in United States. AeroLeads shows a matched LinkedIn profile for Sai Chandra.
Sai Chandra previously worked as Data Engineer at Prudential Financial and Data Engineer at Ditech Group.
Email format at Calpine
This section adds company-level context without repeating Sai Chandra's masked contact details.
Review company-level records connected to Sai Chandra before choosing the right outreach path.
About Sai Chandra
• Over 9+ years of IT experience in a variety of industries working on Big Data technology using technologies such as Cloudera and Hortonworks distributions. Hadoop working environment includes Hadoop, Spark, MapReduce, Kafka, Hive, Ambari, Sqoop, HBase, and Impala. Can work parallelly in both GCP and Azure Clouds coherently. • Hands on experience with different programming languages such as Python, SAS. • Able to use Sqoop to migrate data between RDBMS, NoSQL databases and HDFS. • Experience in Extraction, Transformation and Loading (ETL) data from various sources into Data Warehouses, as well as data processing like collecting, aggregating and moving data from various sources using Apache Flume, Kafka, PowerBI and Microsoft SSIS. • Hands-on experience with Hadoop architecture and various components such as Hadoop File System HDFS, Job Tracker, GCP, Task Tracker, Name Node, Data Node and Hadoop MapReduce programming. • Experience in handling python and spark context when writing Pyspark programs for ETL. • Seasoned practice in Machine Learning algorithms and Predictive Modeling such as Linear Regression, Logistic Regression, Naïve Bayes, Decision Tree, Random Forest, KNN, Neural Networks, and K-means Clustering. • Ample knowledge of data architecture including data ingestion pipeline design, Hadoop/Spark architecture, data modeling, data mining, machine learning and advanced data processing. • Implementations of generalized solution model using AWS SageMaker • Hands-on experience with Amazon EC2, Amazon S3, Amazon RDS, VPC, IAM, Amazon Elastic Load Balancing, Auto Scaling, CloudWatch, SNS, SES, SQS, Lambda, EMR and other services of the AWS family. • Used IDEs like Eclipse, IntelliJ IDE, PyCharm IDE, Notepad ++, and Visual Studio for development. • Very keen in knowing newer techno stack that Google Cloud platform (GCP) adds. • Fluent programming experience with Scala, Python, SQL, T-SQL, R. • Hands-on experience in developing and deploying enterprise-based applications using major Hadoop ecosystem components like MapReduce, YARN, Hive, HBase, Flume, Sqoop, Spark MLlib, Spark GraphX, Spark SQL, Kafka. • Worked on Dimensional Data modelling in Star and Snowflake schemas and Slowly Changing Dimensions (SCD). • Very keen in knowing newer techno stack that Google Cloud platform (GCP) adds. • Experience working with NoSQL databases like Cassandra and HBase and developed real-time read/write access to very large datasets via HBase.
Sai Chandra's current company
Company context helps verify the profile and gives searchers a useful next step.
Sai Chandra work experience
A career timeline built from the work history available for this profile.
Data Engineer
Data Engineer
• Experience in building and architecting multiple Data pipelines, end to end ETL and ELT process for Data ingestion and transformation in GCP and coordinate task among the team. • The roles include creating data pipelines from application databases to Data Lake and warehouse and create stream and batch data processing pipelines. • Build data pipelines in airflow in GCP for ETL related jobs using different airflow operators. • Created python scripts to ingest data from on-premise to GCS and built data pipelines using Apache Beam and Data Flow for data transformation from GCS to Big query • Maintaining more than 16 PB, 300 nodes Cloudera's distribution Hadoop production and dev clusters. Perform daily health checks, work on alerts, and other related tasks. • Created Hive databases, tables and queries for data analytics as and when needed, integrate them with data processing jobs, and drop them as a part of cleanup. Also working on maintaining HBase database used by other applications. • Leveraged Google Cloud Platform Services to process and manage the data from streaming and file-based sources. • Designed and developed in house database visualization tool built on Oracle Application Express platform that visualize the real-time database health information as well as shows management reports. • Experience in moving data between GCP and Azure using Azure Data Factory. • Experienced as AWS cloud engineer to create and manage EC2 instances, S3 buckets, and configure RDS instances. Monitoring the performance, spinning off the instances regularly, and help developers to get access to the cloud instances.
Senior Data Engineer
• Worked on AWS Data pipeline to configure data loads from S3 to into Redshift. • Using AWS Redshift, I Extracted, transformed and loaded data from various heterogeneous data sources and destinations • Created Tables, Stored Procedures, and extracted data using T-SQL for business users whenever required. • Performs data analysis and design, and creates and maintains large, complex logical and physical data models, and metadata repositories using ERWIN and MB MDR • I have written shell script to trigger data Stage jobs. • Assist service developers in finding relevant content in the existing reference models. • Like Access, Excel, CSV, Oracle, flat files using connectors, tasks and transformations provided by AWS Data Pipeline. • Utilized Spark SQL API in PySpark to extract and load data and perform SQL queries. • Worked on developing Pyspark script to encrypting the raw data by using Hashing algorithms concepts on client specified columns. • Responsible for Design, Development, and testing of the database and Developed Stored Procedures, Views, and Triggers • Developed Python-based API (RESTful Web Service) to track revenue and perform revenue analysis. • Compiling and validating data from all departments and Presenting to Director Operation. • KPI calculator Sheet and maintain that sheet within SharePoint.
Data Engineer
• Designed and Implemented Sharding and Indexing Strategies for MongoDB servers. • Optimizing pig scripts, user interface analysis, performance tuning and analysis. • Used Hive to analyze the partitioned and bucketed data and compute various metrics for reporting on the dashboard. • Implemented the import and export of data using XML and SSIS. • Involved in Planning, Defining and Designing data base using ER Studio on business requirement and provided documentation. • Responsible for migration of application running on premise onto Azure cloud. • Used SSIS to build automated multi-dimensional cubes. • Wrote indexing and data distribution strategies optimized for sub-second query response • Developed a statistical model using artificial neural networks for ranking the students to better assist the admission process. • Wrote scripts and indexing strategy for a migration to Confidential Redshift from SQL Server and MySQL databases. • Implement software enhancements to port legacy software systems to Spark and Hadoop ecosystems on Azure Cloud.
Data Engineer
• Deployed Lambda and other dependencies into AWS to automate EMR Spin for Data Lake jobs • Scheduled spark applications/Steps in AWS EMR cluster. • Extensively used event-driven and scheduled AWS Lambda functions to trigger various AWS resources. • Implemented advanced procedures like text analytics and processing using teh in-memory computing capabilities like Apache Spark written in Scala. • Developed spark applications for performing large scale transformations and denormalization of relational datasets. • Developed and executed a migration strategy to move Data Warehouse from SAP to AWS Redshift. • Loaded data into teh cluster from dynamically generated files using Flume and from relational database management systems using Sqoop. • Used Spark Streaming to divide streaming data into batches as an input to spark engine for batch processing. • Worked on analyzing Hadoop cluster and different Big Data analytic tools including Pig, hive, HBase, Spark and Sqoop. • Exported data from HDFS to RDBMS via Sqoop for Business Intelligence, visualization, and user report generation. • Loading teh data from multiple Data sources like (SQL, DB2, and Oracle) into HDFS using Sqoop and load into Hive tables. • Performed Real time event processing of data from multiple servers in teh organization using Apache Storm by integrating wif apache Kafka. • Performed Impact Analysis of teh changes done to teh existing mappings and provided teh feedback • Involved in complete Implementation lifecycle, specialized in writing custom MapReduce, and Hive • Extensively used Hive/HQL or Hive queries to query or search for a string in Hive tables in HDFS
Data Analyst
Involved in implementation of the project went through several phases namely: data set analysis, preprocessing data set, user-generated data extraction, and modeling.Participated in Data Acquisition with the Data Engineer team to extract historical and real-time data by using Sqoop, Pig, Flume, Hive, MapReduce, Land HDES.Wrote user-defined functions (UDFs) in Hive to manipulate strings, dates, and other data.Performed Data Cleaning, features scaling, features engineering using pandas, and NumPy packages in python .Process Improvement: Analyzed error data of recurrent programs using Python and devised a new process to reduce the turnaround time of the problem's solutions by 60%Worked on production data fixes by creating and testing SQL scripts.Deep dived into complex data sets to analyze trends using Linear Regression, Logistic Regression, Decision Trees.Prepared reports using SQL and Excel to track the performance of websites and apps.Performed Data Collection, Data Cleaning, Data Visualization, and Feature Engineering using Python libraries such as Pandas, Numpy, MatPlotLib, and seaborn.
Colleagues at Calpine
Other employees you can reach at calpine.com. View company contacts for 2309 employees →
Eulete Patrick
Colleague at CalpineHouston, Texas, United States
View →
KP
Katherine Potter
Colleague at CalpineSan Jose, California, United States
View →
AM
Adam Miller
Colleague at CalpineKelseyville, California, United States
View →
BN
Ben Noel
Colleague at CalpineHouston, Texas, United States
View →
PO
Pherone Oates
Colleague at CalpineOakland, California, United States
View →
ST
Sharla Tipton
Colleague at CalpineMagnolia, Texas, United States
View →
JS
Jessie Seholm
Colleague at CalpineKennewick, Washington, United States
View →
MW
Michael Welch
Colleague at CalpineSan Antonio, Texas Metropolitan Area, United States
View →
ZQ
Zhuo Qi
Colleague at CalpineHouston, Texas, United States
View →
PC
Pablo Chaves
Colleague at CalpineSeminole, Texas, United States
View →
Frequently asked questions about Sai Chandra
Quick answers generated from the profile data available on this page.
What company does Sai Chandra work for?
Sai Chandra works for Calpine.
What is Sai Chandra's role at Calpine?
Sai Chandra is listed as Senior Data Engineer at Calpine.
Where is Sai Chandra based?
Sai Chandra is based in United States while working with Calpine.
What companies has Sai Chandra worked for?
Sai Chandra has worked for Calpine, Prudential Financial, Ditech Group, Chewy, and Macy'S.
Who are Sai Chandra's colleagues at Calpine?
Sai Chandra's colleagues at Calpine include Eulete Patrick, Katherine Potter, Adam Miller, Ben Noel, and Pherone Oates.
How can I contact Sai Chandra?
You can use AeroLeads to view verified contact signals for Sai Chandra at Calpine, including work email, phone, and LinkedIn data when available.
Search by job title, company, industry, location, and seniority. Export verified B2B contact data when you need it.
Start free trialCheck these profiles if this is not the Sai Chandra you were looking for.
View similar profiles