Shubham Pal

Shubham Pal

Lead Data Scientist @ AT&T

About

• Data scientist in executing data-driven solutions with adept knowledge on Data Analytics, Text Mining, Machine Learning (ML), Predictive Modeling, and Natural Language Processing(NLP). • Strong mathematical background in Machine Learning, Predictive Modeling and Data Mining with a broad understanding of Supervised and Unsupervised learning techniques and algorithms (eg: Regression, k-NN, Random Forest, SVM, Naïve Bayes, Decision trees, Clustering, etc.) • Solid experience in Deep Learning techniques with Convolutional Neural Networks (CNN), Recursive Neural Networks (RNN), max pooling and normalization. • Extracted data from various source systems like Oracle, SQL Server and flat files as per the requirements. • Highly competent at wide varieties of Data Science programming languages such as Python, R, SQL and visualization in Tableau. • Involved in entire data science project life cycle, including Data Acquisition, Data Cleansing, Data Manipulation, Feature Engineering, Modeling, Evaluation, Optimization, Testing and Deployment.

Country

United States

City

New York City Metropolitan Area

Industry

Information Technology & Services

Skill

Financial Analysis, Data Analysis, jQuery, Java, Project Management, Statistics, Machine Learning, Python, Matlab, SQL, Microsoft Office, Microsoft Excel, Microsoft Word, Microsoft PowerPoint, R , PostgreSQL, Tableau, Leadership, Public Speaking, Teamwork

Experience

AT&T

Lead Data Scientist

AT&T

LinkedIn
2024-1 - Present · 2 yrs 9 mos

1. Leading Finance Transformation by deploying agentic AI systems (LangGraph, deep agents) to automate financial analysis, forecasting, and decision-making at scale. 2. Architected modular multi-agent workflows (EDA, modeling, evaluation) with dynamic orchestration, significantly improving speed and accuracy of financial insights. 3. Built and fine-tuned GenAI + RAG systems (PEFT/LoRA) for audit reporting and policy intelligence, reducing manual effort and standardizing outputs. 4. Enabled AI-powered analytics in Power BI, delivering real-time, query-driven insights for business leaders. Developed enterprise-scale cash flow & disbursement forecasting models used in executive planning and investor guidance. 5. Drove strategic decision-making on capital allocation and liquidity, contributing to multi-year stock growth and ~$75B increase in market cap through high-confidence forecasts. 6. Built Auto-Taxability ML system (AutoML + NLP/transformers) to classify tax applicability at scale. 7. Delivered ~$1M/month savings by eliminating overpayments and improving compliance. 8. Productionized ML systems using Python, PySpark, Azure Databricks, handling large-scale financial data pipelines. 9. Mentored teams and scaled adoption of agentic AI and AI-first engineering practices across org.

AT&T

Senior Data Scientist

AT&T

LinkedIn
2022-4 - 2024-1 · 1 yr 10 mos
AT&T

Data Scientist

AT&T

LinkedIn
2021-6 - 2022-4 · 11 mos
New York State Office of Mental Health

Data Scientist

New York State Office of Mental Health

LinkedIn
2020-7 - 2021-5 · 11 mos

Albany, New York, United States

• Extracted, processed big datasets and designed deep learning models like GRU, LSTMs in python for time series forecasting of various rare events like hospitalization, FEP (First Episode of psychosis) for early-stage intervention program. • Used SMOTE, class weights and temporal sample weights for dealing with highly imbalanced class data problem. • Develop, design and deploy various machine learning algorithms like regression, negative binomial, decision tress, XGBoost for feature extraction and predictive analysis of clinical data. • Build text analysis pipeline using NLP, Word2Vec and BERT to cluster manually entered string records into meaningful groups. • Used SAS for basic statistical analysis of Medicaid data and other ad hoc requests. • Wrote PL/SQL queries to pull relevant data from MDW (Medicaid Data Warehouse) and monthly update tables and server migration.

New York State Office of Mental Health

Data Analyst

New York State Office of Mental Health

LinkedIn
2020-2 - 2020-6 · 5 mos

Albany, New York Area

• Extract, interpret and analyze data to identify key metrics and transform large raw data into relevant, actionable information using PL/SQL & SAS. • Gather data from MDW (Medicaid Data Warehouse) using Oracle SQL Developer according to data requests, Developed complex SQL queries using Window Functions , CTE and Optimized for performance using Views. • Analyzed network adequacy and impact of various healthcare service providers. • Collect, cleanse and provide modeling and analyses of structured and unstructured data for Office of Mental Health. • Wrangling ,cleaning and merging big datasets using Pyspark and Python Dask for creating a feeder database for tableau dashboards. • Create visually impactful dashboards in Tableau for data reporting using roll-up tables and publish them on server, also developed similar visualizations using Python Matplotlib and Seaborn. • Create reports to share findings and recommendations with the internal team and other stakeholders.

NYSERDA

Data Analyst

NYSERDA

LinkedIn
2019-4 - 2020-2 · 11 mos

Albany, New York Area

-Engineered real-time financial reports & dashboards in Tableau for NY’s residential and commercial projects. -Spearheaded multiple analyses, devising predictive models & algorithms on a wide range of key metrics using SAS ,SQL and Python. -Presented insights & strategy recommendations to project managers, team leads, & executive management. -Manipulated big data using PySpark and Dask to provide queries, statistical summaries, and data visualizations. -Forecasted consumption / engagement & identified consumer preferences to augment social media outreach. -Built Predictive models for customer segmentation and power consumption cost analysis to understand customer energy consumption behaviour using Sklearn, Python and R.

Matrixpro

Data Scientist / Engineer

Matrixpro

LinkedIn
2017-5 - 2018-4 · 1 yr

* Developed a machine learning system that predicted purchase probability at a particular offer based on customer’s real time location data and past purchase behaviour. • Performed k-Means clustering in order to understand customer itemized bought products and segment the customers based on the customer products for animal medicine and vaccines behaviour information for customized product offering, customized and priority service, to improve existing profitable relationships and to avoid customer churn, etc. using Python. • Used PySpark Machine learning library to build and evaluate different models. • Generated, wrote and ran SQL script to implement the database changes including table update, addition or update of indexes, creation of views and store procedures. • Developed advanced SQL queries with multi table joins, groups, functions, sub queries, set operations, & stored procedures for Data Analysis in SQL server. • Programmed ETL functions between Oracle and Amazon Redshift. Managed big data files in AWS S3 using boto3 client. • Developed and maintained stored procedures and complex packages extensively using PL/SQL and shell programs.

Education

The State University of New York

The State University of New York

LinkedIn

Data science

Shubham Pal's Contact Information

Email

******@***.com

Phone

(**) *** ****

Find the Right Leads
Find Verified Contact Data

Try with: Jensen Huang @ nvidia.com Click to autofill
LeadContact awards, five-star ratings, and GDPR compliance badges

What LeadContact does well

Find verified emails, phone numbers, and decision-makers with 98% accuracy.

Find Leads

Find Leads

Find the right people by company, role, industry, location, and more.

925M+ professional profiles

Find Leads
Find Emails

Find Emails

Access verified email addresses for your target contacts.

657M+ emails

Find Emails
Find Phone Numbers

Find Phone Numbers

Get cross-validated phone data from multiple top sources.

239M+ phone numbers

Find Phone Numbers

More Accurate. Lower Cost.

Find contact data in 1 tool with 98% accuracy

LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.

LeadContact Logo
Competitor Tools

All these = $289 per month

Great conversations start with the right contact.

It’s time to find yours.