Vino Duraisamy

Vino Duraisamy

Senior Developer Advocate - Data & AI @ Snowflake

About

📌 "In God we trust; all others must bring data." Data & AI practitioner with a unique blend of data engineering, machine learning, artificial intelligence and business analysis expertise. Currently, as a Developer advocate at Snowflake I work on Data Engineering and AI workloads. In my professional experience, I have worked on end-to-end data & AI projects that involved Data Engineering, Data Modeling, Machine Learning Model Deployment, Data Visualization, and Analytics Framework Development for solving business problems. Most recently, as a Data Engineer at Applied ML Search team at Apple, I built a search metrics data store for retail/e-commerce data, and developed automated training pipelines for search query intent classification model and learn-to-rank ML models in production. Previously, as an AI Engineer at IproTech, I worked on a custom NLP model for Named Entity Recognition that identifies and redacts personally identifiable and sensitive information from unstructured text documents. Prior to that, I worked as a Data Engineer at Nike creating robust and scalable enterprise data pipelines for the Consumer Data Analytics & Engineering (CDEA) team. ✅ Programming: Python (Numpy, Pandas, NLTK, Spacy, Scikit-Learn), PySpark, SQL. ✅ Big Data Tools & Frameworks: Apache Spark, Snowflake, Hadoop, ETL data pipelines, Apache Airflow, Hive, AWS (EMR, S3, EC2, Lambda, SNS, Glue, Redshift), lakeFS ✅ VCS, DevOps & Misc Tools: PyCharm, VS Code, Git, Jira, Jenkins, Docker ✅ Statistics: Inferential Statistics, Experimental Design, Hypothesis Testing (A/B Testing), Regression Analysis ✅ Machine Learning: Regression Modeling, Random Forest, XGBoost, kNN Classifier, K-means Clustering, Feature Extraction (PCA, Factor Analysis), Natural Language Processing (Text Analytics – PII, PHI Extraction), Convolutional Neural Network. ✅ Business Domain Expertise: Enterprise Data Analytics (Data Warehousing), Sales and Marketing Analytics, Customer Segmentation, Customer Lifetime Value and Retention Analysis, Customer Success KPIs, Product Analytics

Country

United States

City

San Francisco

Industry

Computer Software

Skill

Artificial Intelligence (AI), High Performance Organizations, Snowflake, Machine Learning, Python (Programming Language), MySQL, Data Engineering, lakeFS, Convolutional Neural Networks (CNN), Keras, NumPy, Matplotlib, Apache Oozie, Docker, OpenAI, Jenkins, Pandas, Scikit-Learn, NLTK, Data Analysis

Experience

Snowflake

Senior Developer Advocate - Data & AI

Snowflake

LinkedIn
2025-8 - Present · 1 yr 2 mos
Snowflake

Developer Advocate - Data & AI

Snowflake

LinkedIn
2023-7 - 2025-8 · 2 yrs 2 mos

San Francisco Bay Area

Towards Data Science

Data Engineering Writer

Towards Data Science

LinkedIn
2023-1 - Present · 3 yrs 9 mos
Medium

Technical Writer - Data Engineering, AI, ML, Statistics

Medium

LinkedIn
2020-5 - Present · 6 yrs 5 mos

● Publishing articles on a variety of topics including Data Engineering, AI, ML, Statistics and Data Analytics.

lakeFS

Developer Advocate - Data Engineering, Machine Learning, MLOps

lakeFS

LinkedIn
2022-5 - 2023-6 · 1 yr 2 mos

San Francisco Bay Area

lakeFS is an open-source tool that offers data versioning at scale for object stores - think "git for data", only it scales for 100s of petabytes datalakes. ● Start managing data the way you manage your code. ● lakeFS sits on top of your object store and provides git-like capabilities such as commit, branch, revert or merge all via UI, CLI or API. ● The ecosystem of tools you have today can access the data via lakeFS the same way they access it today. Easy peasy! Yep, I said it.

Apple

Data Engineer - Applied AI & Search

Apple

LinkedIn
2021-4 - 2022-5 · 1 yr 2 mos

Sunnyvale, California, United States

reveal - ipro

AI Engineer - Language Models (NLP)

reveal - ipro

LinkedIn
2020-9 - 2021-3 · 7 mos

Tempe, Arizona, United States

Tools & Languages Used: Python (Spacy), Named Entity Recognition Models (Decision Tree, Random Forest, CNN), Docker, Jenkins ● Developed and trained a Spacy based Named Entity Recognition model to identify and redact personally identifiable information from unstructured text data. ● Preprocessed and cleaned a corpus of 1M+ documents (emails, tweets, wikipedia and news articles) using Python. ● Engineered features for different document types and built an integrated feature pipeline for model training. ● Achieve an average F1-score of 80% for PII data and deployed the model as a python package in production. ● Worked with product managers and development teams to ensure accurate integration of the model into the product.

C1X Inc.

Data Engineer (client: Nike)

C1X Inc.

LinkedIn
2020-8 - 2020-10 · 3 mos

San Jose, California, United States

Tools & Languages Used: Python, Spark, Hive, AWS (S3, EMR, DynamoDB), Airflow, Hadoop, Docker, Jenkins ● Developed robust, scalable data pipelines for data privacy compliance module in Airflow for Customer Data Engineering & Analytics team at Nike. ● Created several DAGs to ensure hourly ingestion of customer data from AWS S3 buckets into Hive tables. ● Designed and built re-usable libraries in PySpark to ensure data quality and to support data engineering, and downstream analytics workflows. ● Continuously monitored and improved data pipelines by analyzing bottlenecks, and implemented efficient solutions. ● Refactoring & optimizing the code for improved reliability, performance, simplicity and maintenance.

reveal - ipro

AI Engineer - Language Models (NLP)

reveal - ipro

LinkedIn
2019-11 - 2020-5 · 7 mos

Phoenix, Arizona Area

Tools used: Python (SpaCy, Gensim, Regex), Prodigy, AWS S3 & EC2 ● Developed a prototype PII (Personally Identifiable Information) extraction and Named Entity Recognition model for Email corpus using Python. ● Leveraged SpaCy, Gensim and regex libraries for data cleaning and pre-processing to improve data quality. ● Built SpaCy based NER model and rule based PII extraction engine to identify, tag and redact PII data including health, finance, security and other sensitive information (such as Social Security Number, Bank Account Number, Credit Card number, etc. ) ● Re-trained SpaCy English language model ‘en_core_web_lg’ on Email corpus using the data annotation and training tool Prodigy for improved model performance.

Arizona State University

Explainable AI Researcher

Arizona State University

LinkedIn
2019-9 - 2020-5 · 9 mos

Tempe, Arizona, United States

● Worked with Dr. Asim Roy from Dept of Information Systems on improving explainability and interpretability of deep learning models. ● Built a 4 layer Convolutional Neural Network for MNIST handwritten digit recognition with 99.9% accuracy as the base model for analysis. ● Analyzed the filters, pooling layers and inter-connected layers of convolutional neural networks and visualized the activation values at each layer to understand underlying abstractions at each level. ● Leveraged Activation maximization to generate the input image that maximizes activations of particular neuron or a group of neurons, thus revealing the abstractions captured by the neurons. ● Used Saliency maps to understand the part of input image that contributes to the activations.

Adrenalin.hr

Analytics/ Data Engineer - Sales & Marketing

Adrenalin.hr

LinkedIn
2018-10 - 2019-6 · 9 mos

Chennai Area, India

Tools used: Python, Tableau, MySQL, Excel ● Gathered data on customers, competitive products and market place, and consolidate information into actionable items, reports and presentations. ● Built a segmentation model to classify prospective customers using different factors enabling sales teams to derive efficient strategies for different prospect groups, thus decreasing the average sales cycle. ● Developed a customer churn prediction model to identify the top 3 factors contributing to customer churn, enabling the account managers to work on retention strategy. This increased the retention rate and improved the overall customer satisfaction score as well. ● Created Tableau dashboards and periodic reports for senior leadership to track customer churn metrics.

Education

W. P. Carey School of Business – Arizona State University

W. P. Carey School of Business – Arizona State University

LinkedIn

Business Analytics

PSG College of Technology

PSG College of Technology

LinkedIn

Electronics and Communications Engineering

Vino Duraisamy's Contact Information

Email

******@***.com

Phone

(**) *** ****

Find the Right Leads
Find Verified Contact Data

Try with: Jensen Huang @ nvidia.com Click to autofill
LeadContact awards, five-star ratings, and GDPR compliance badges

What LeadContact does well

Find verified emails, phone numbers, and decision-makers with 98% accuracy.

Find Leads

Find Leads

Find the right people by company, role, industry, location, and more.

925M+ professional profiles

Find Leads
Find Emails

Find Emails

Access verified email addresses for your target contacts.

657M+ emails

Find Emails
Find Phone Numbers

Find Phone Numbers

Get cross-validated phone data from multiple top sources.

239M+ phone numbers

Find Phone Numbers

More Accurate. Lower Cost.

Find contact data in 1 tool with 98% accuracy

LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.

LeadContact Logo
Competitor Tools

All these = $289 per month

Great conversations start with the right contact.

It’s time to find yours.