Anurag Hambir
Data Scientist @ Credit One Bank
About
I am a Data Scientist with over 6 years of experience in Data Science and engineering. I recently completed my Master's degree in Data Science at Indiana University Bloomington in May 2023 where I learned advanced skills in machine learning, data visualization, and cloud computing. Across my professional career, I have managed and implemented various data science projects involving large structured and unstructured data sets, using tools and techniques such as Python, PySpark, SQL, NoSQL, Kafka, Graph Neural Networks, and Large Language Models. I enjoy working with diverse teams and solving complex problems that have a positive impact on the society. I am eager to apply my skills and knowledge to new challenges and opportunities, and to contribute to the innovation and growth of the data science field.
United States
Las Vegas
Information Technology & Services
Large Language Models (LLM), Object-Oriented Programming (OOP), Data Structures, Algorithm Design, Elasticsearch, Artificial Intelligence (AI), Natural Language Processing (NLP), PySpark, Hive, Amazon Web Services (AWS), PyTorch, XGBoost, Big Data , Recommender Systems, Data Engineering, Cloud Computing, Statistics, Big Data Analytics, PostgreSQL, Data Science
Experience

Data Scientist
Social Science Consulting LLC
Bloomington, Indiana, United States
• Employed advanced large language models (BERT and RoBERTa) in conjunction with PyTorch (Python) on a 300,000-text data set to create precise contextual word embeddings for generating personalized recommendations • Designing and leading the development of Graph Neural Network-based Recommendation system development (GraphSAGE, GATConv), enabling personalized foundation suggestions for grant recipients, and promoting reciprocal recommendations

Data Scientist
United States
• Worked on the development of a customer retention model using Python for a Fortune 500 banking and insurance client, boosting the F-1 score by 70% through the XGBoost model and genetic algorithm, resulting in projected savings of $269,000 from retained customers • Utilized SQL and Snowflake to add new features to the Attrition model • Optimized and functionalized LSTM model for predicting the percentage of clients who churn

Data Scientist
India
• Managed various data projects involving large unstructured data sets, while offering guidance and mentorship to a team of engineers. • Designed and implemented ETL pipelines for crawling data from different sources and stored them in MongoDB(NoSQL) database. • Built and deployed a recommendation engine for the OTTPlay application.

Data Engineer
Pune, Maharashtra
• Implemented a data pipeline using PySpark to extract and analyze sentiments and entities from clients’ email data, employing Stanford’s CoreNLP library and SpaCy. • Processed and analyzed email data stored in Hadoop, through PySpark and stored the results in Hive tables. • Improved the performance of the data pipeline by integrating Kafka, reducing the total processing time by 50%.
Anurag Hambir's Contact Information
Phone
Find the Right Leads
Find Verified Contact Data
What LeadContact does well
Find verified emails, phone numbers, and decision-makers with 98% accuracy.
Find Leads
Find the right people by company, role, industry, location, and more.
925M+ professional profiles

Find Emails
Access verified email addresses for your target contacts.
657M+ emails

Find Phone Numbers
Get cross-validated phone data from multiple top sources.
239M+ phone numbers

More Accurate. Lower Cost.
Find contact data in 1 tool with 98% accuracy
LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.
Great conversations start with the right contact.
It’s time to find yours.




