AeroLeads people directory · profile

G D. Email & Phone Number

Senior Data Scientist at Capital One
Location: Charlotte, North Carolina, United States 8 work roles
1 work email found @capitalone.com LinkedIn matched
✓ Verified August 2026 3 data sources Profile completeness 86%

Contact Signals · 1 work email

Work email g****@capitalone.com
LinkedIn Profile matched
3 free lookups remaining · No credit card
Current company
Role
Senior Data Scientist
Location
Charlotte, North Carolina, United States
Company size

Who is G D.? Overview

A concise factual answer block for searchers comparing this professional profile.

Quick answer

G D. is listed as Senior Data Scientist at Capital One, a with 63917 employees, based in Charlotte, North Carolina, United States. AeroLeads shows a work email signal at capitalone.com and a matched LinkedIn profile for G D..

G D. previously worked as Lead Data Scientist at Marketamerica and Senior Data Scientist at Ford Motor Company.

Company email context

Email format at Capital One

This section adds company-level context without repeating G D.'s masked contact details.

*@capitalone.com
71% confidence

AeroLeads found 1 current-domain work email signal for G D.. Compare company email patterns before reaching out.

Profile bio

About G D.

• A highly analytical and productive professional with over 10+ years of experience in Data Science. Excellent in developing an engagement plan using the technical expertise and business acumen to translate the business requirements and project objectives into successful system solutions.• Expertise on designing and developing the Big data Analytics platforms for Retail, Logistics, Healthcare and Banking Industries using Big Data, Spark, Real-time streaming, Kafka, Data Science, Machine Learning, NLP and Cloud• Experience in end-to-end Data Development project hands-on starting from Data Ingestion, Data Quality, Data Governance, Data Management, Data Loading, Data Reporting and Analyzing.

Listed skills include Microsoft Office, Strategic Planning, Business Development, Microsoft Excel, and 24 others.

Current workplace

G D.'s current company

Company context helps verify the profile and gives searchers a useful next step.

Capital One
Capital One
Senior Data Scientist
Charlotte, NC, US
Website
Employees
63917
AeroLeads page
8 roles

G D. work experience

A career timeline built from the work history available for this profile.

Lead Data Scientist

Current

• Spearheaded NLP project for e-commerce categorization, leveraging LLM and GPT-4 models, enhancing product classification and search relevance, driving revenue growth.• Led development of vector search system using advanced NLP, including GPT-4, boosting product discovery and conversion rates.• Collaborated across teams to optimize machine learning algorithms using quantization methods and perturbation techniques, aligning with business goals for revenue generation.• Developed personalized recommendation models using LLM and NLP, incorporating customer data for improved satisfaction and driving repeat purchases.• Deployed ML models, including LLM, on AWS SageMaker for scalability and real-time inference, integrating seamlessly with existing infrastructure.• Optimized ML models through continuous experimentation, including perturbation techniques, staying ahead in NLP and recommendation systems.• Provided expertise in NLP and ML, fostering innovation and knowledge sharing within internal teams, enhancing AI capabilities.

Sep 2022 - Present

Senior Data Scientist

Dearborn, Michigan, Us

• Customer Segmentation: Built a customer segmentation model using regression models in Python to segment millions of customers to gain insight into behaviors, measure marketing effectiveness, and better allocate future marketing spend. • Customer Churn: Built and maintained logistic regression and random forest models which predicted the customer’s likelihood to renew vs account closures with 85-90% accuracy ($20M balance saved) based on customer usage patterns and characteristics• Lead in initiative to build statistical models using historical data and consumer data to identify the potential consumers for the cross selling financial products • Construct and fit statistical, machine learning, or optimization models that enable estimation of retail establishment survey decision-making across a range of complex environments and applications• Used pandas, NumPy, seaborne, SciPy, Matplotlib, Scikit-learn, NLTK in Python for developing various machine learning algorithms. Expertise in R, Mat lab, python and respective libraries.• Performed K-means clustering, Regression and Decision Trees in Python. Worked on data cleaning and reshaping, generated segmented subsets using NumPy and Pandas in Python.• Identified and removed outliers in the data by using different statistical methods like Standard Deviation Method and Inter Quartile Range (IQR) Methods• Tackled highly imbalanced Fraud dataset using under sampling, oversampling with SMOTE and cost sensitive algorithms with Python Scikit-learn.• Hands on experience in implementing Naive Bayes and skilled in Random Forests, Decision Trees, Linear, and Logistic Regression, Clustering,• Power BI development and administration.

Jul 2021 - Aug 2022

Lead Data Scientist

Mclean, Va, Us

• Wrote complex Spark SQL queries for data analysis to meet business requirement.• Participated in feature engineering such as feature intersection generating, feature normalize and label encoding with Scikit-learn preprocessing.• Improved fraud prediction performance by using random forest and gradient boosting for feature selection with Python Scikit-learn.• Used big data tools Spark (Pyspark, SparkSQL, Mllib) to conduct real time analysis of loan default based on AWS.• Conducted Data blending, Data preparation using Alteryx and SQL for tableau consumption and publishing data sources to Tableau server.• Assist customers by being able to deliver a ML project from beginning to end, including understanding the business need, aggregating data, exploring data, building & validating predictive models, and deploying completed models with concept-drift monitoring and retraining to deliver business impact to the organization • Use AWS AI services (e.g., Personalize), ML platforms (SageMaker), and frameworks (e.g.,TensorFlow, PyTorch, SparkML, scikit-learn) to help our customers build ML models

Jul 2020 - Mar 2021

Senior Data Scientist

New York, Ny, Us

• Using NLP, extract key information from medical reports thereby reducing the processing time for more standard cases and enabling underwriters to focus on the most difficult or complex ones.• Created Pipelines in ADF using Linked Services/Datasets/Pipeline/ to Extract, Transform and load data from different sources like Azure SQL, Blob storage, Azure SQL Data warehouse, write-back tool and backwards.• Using NLP techniques (with NLTK and gensim libraries), prototyped automatic extraction of structured listings data from free-form text descriptions. • Use AWS AI services (e.g., Personalize), ML platforms (SageMaker), and frameworks (e.g., MXNet, TensorFlow, PyTorch, SparkML, scikit-learn) to help our customers build ML models• Apply computer based mathematical/statistical techniques using software. Lead or participate in statistical projects or studies in survey sampling (design and estimation), modeling, or statistical research. Apply knowledge of programming/coding language (i.e., SQL, SAS, SPSS, RStudio) to develop scripts or applications• Research and implement novel ML approaches, including hardware optimizations on platforms such as AWS Inferentia • Collaborated with product team to define KPIs and to assess the progress thereof; also, to propose and execute product analytics projects such as user segmentation• Identified and removed outliers in the data by using different statistical methods like Standard Deviation Method and Inter Quartile Range (IQR) Methods• Handled imbalanced data sets by resampling methods like Synthetic Minority over Sampling Technique (SMOTE) and Random under Sampling methods• Tackled highly imbalanced Fraud dataset using under sampling, oversampling with SMOTE and cost sensitive algorithms with Python Scikit-learn.• Wrote complex Spark SQL queries for data analysis to meet business requirement.• Developed MapReduce/Spark Python modules for predictive analytics & machine learning in Hadoop on AWS.

Apr 2016 - Dec 2019

Data Scientist

Smart Tricks

• Lead in initiative to build statistical models using historical data to predict FMCG sales in several economic markets. Focused on analyzing the factors affecting the sales of SENAP region• Construct and fit statistical, machine learning, or optimization models that enable estimation of retail establishment survey decision-making across a range of complex environments and applications• Used pandas, NumPy, seaborne, SciPy, Matplotlib, Scikit-learn, NLTK in Python for developing various machine learning algorithms. Expertise in R, Mat lab, python and respective libraries.• Research on Reinforcement Learning and control (Tensor Flow, Torch), and machine learning model (Scikit-learn).• Hands on experience in implementing Naive Bayes and skilled in Random Forests, Decision Trees, Linear, and Logistic Regression, SVM, Clustering, Principal Component Analysis.• Performed K-means clustering, Regression and Decision Trees in R. Worked on data cleaning and reshaping, generated segmented subsets using NumPy and Pandas in Python.• Implemented various statistical techniques to manipulate the data like missing data imputation, principal component analysis and sampling.• Worked on R packages to interface with Caffe Deep Learning Framework. Perform validation on machine learning output from R.• Applied different dimensionality reduction techniques like principal component analysis (PCA) and t-stochastic neighborhood embedding (t-SNE) on feature matrix.• Performed univariate and multivariate analysis on the data to identify any underlying pattern in the data and associations between the variables.• Responsible for design and development of Python programs/scripts to prepare transform and harmonize data sets in preparation for modeling.• Worked with Market Mix Modeling to strategize the advertisement investments to better balance the ROI on advertisements.

Apr 2015 - Apr 2016

Data Analyst

Medrcedu Tech

• Installed and configured Apache Hadoop to test the maintenance of log files in Hadoop cluster.• Involved in Requirement gathering, Business Analysis and translated business requirements into Technical design in Hadoop and Big Data• Involved in SQOOP implementation which helps in loading data from various RDBMS sources to Hadoop systems and vice versa. • Developed Python scripts to extract the data from the web server output files to load into HDFS.• Involved in HBASE setup and storing data into HBASE, which will be used for further analysis.• Worked on Cloud Health tool to generate AWS reports and dashboards for cost analysis.• Written a python script which automates to launch the EMR cluster and configures the Hadoop applications.• Extensively worked with Avro and Parquet files and converted the data from either format Parsed Semi Structured JSON data and converted to Parquet using Data Frames in PySpark.• Experienced in analyzing and Optimizing RDD's by controlling partitions for the given data• Experienced in writing live Real-time Processing using Spark Streaming with Kafka• Used HiveQL to analyze the partitioned and bucketed data and compute various metrics for reporting• Experienced in querying data using SparkSQL on top of Spark engine• Involved in managing and monitoring Hadoop cluster using Cloudera Manager.• Used Python and Shell scripting to build pipelines.• Developed data pipeline using sqoop, HQL, Spark and Kafka to ingest Enterprise message delivery data into HDFS.• Plan, develop, coordinate, and participate in various marketing research activities to identify customer preferences and attitudes and to enhance products and services• Created data partitions on large data sets in S3 and DDL on partitioned data• Worked for BI Analytics team to conduct A/B testing, data extraction and exploratory analysis.

May 2013 - Apr 2015

Data Analyst

Hyderabad, Telangana, In

• Research and recommend suitable technology stack for Hadoop migration considering current enterprise architecture.• Responsible for building scalable distributed data solutions using Hadoop.• Experienced in loading and transforming of large sets of structured, semi-structured and unstructured data.• Developed Spark jobs and Hive Jobs to summarize and transform data.• Built on-premises data pipelines using Kafka and spark for real-time data analysis.• Created reports in TABLEAU for visualization of the data sets created and tested Spark SQL connectors.• Implemented Hive complex UDF's to execute business logic with Hive Queries.• Developed a different kind of custom filters and handled pre-defined filters on HBase data using API.• Implemented Spark using Scala and utilizing Data frames and Spark SQL API for faster processing of data.• Handled importing data from different data sources into HDFS using Sqoop and performing transformations using Hive and then loading data into HDFS.• Exporting of a result set from HIVE to MySQL using Sqoop export tool for further processing.• Collecting and aggregating large amounts of log data and staging data in HDFS for further analysis.• Experience in managing and reviewing Hadoop Log files.• Experienced in querying data using SparkSQL on top of Spark engine• Involved in managing and monitoring Hadoop cluster using Cloudera Manager.• Used Python and Shell scripting to build pipelines.• Developed data pipeline using sqoop, HQL, Spark and Kafka to ingest Enterprise message delivery data into HDFS.• Used Sqoop to channel data from different sources of HDFS and RDBMS.• Developed Spark applications using Pyspark and Spark-SQL for data extraction, transformation and aggregation from multiple file formats.• Created Partitioned Hive tables and worked on them using HiveQL.• Loading Data into HBase using Bulk Load and Non-bulk load

Jun 2008 - Apr 2010
Team & coworkers

Colleagues at Capital One

Other employees you can reach at capitalone.com. View company contacts for 63917 employees →

FAQ

Frequently asked questions about G D.

Quick answers generated from the profile data available on this page.

What company does G D. work for?

G D. works for Capital One.

What is G D.'s role at Capital One?

G D. is listed as Senior Data Scientist at Capital One.

What is G D.'s email address?

AeroLeads has found 1 work email signal at @capitalone.com for G D. at Capital One.

Where is G D. based?

G D. is based in Charlotte, North Carolina, United States while working with Capital One.

What companies has G D. worked for?

G D. has worked for Capital One, Marketamerica, Ford Motor Company, Nielsen, and Smart Tricks.

Who are G D.'s colleagues at Capital One?

G D.'s colleagues at Capital One include Marquinn Greer, Diana Cunningham, Cams, Jesus Hinojosa, Celeste Joy Llanes, and Ariel Howard.

How can I contact G D.?

You can use AeroLeads to view verified contact signals for G D. at Capital One, including work email, phone, and LinkedIn data when available.

What skills is G D. known for?

G D. is listed with skills including Microsoft Office, Strategic Planning, Business Development, Microsoft Excel, Crm, Business Strategy, Pre Sales, and Business Analysis.

Find 750M verified contacts

Search by job title, company, industry, location, and seniority. Export verified B2B contact data when you need it.