
Harshith G
AI/ML Engineer @ Stripe
About
I’m an AI/ML Engineer specializing in LLM systems, RAG architectures, agentic workflows, and scalable ML pipelines. I focus on building production-ready AI systems that combine retrieval, reasoning, and automation to solve real business problems. At Stripe, I’ve built LLM-powered internal assistants, hybrid retrieval pipelines, and multi-agent workflows using LangChain, LangGraph, Pinecone, FAISS, and FastAPI. I design modular AI architectures involving embedding services, vector indexing, hybrid retrieval, re-ranking, context assembly, and GPU-backed inference. My work has improved document relevance, reduced latency, and helped automate high-volume dispute and support workflows. Before that, at Mastercard, I developed fraud detection models, built PySpark + Kafka streaming pipelines, and helped design components of real-time ML decisioning systems. This gave me a strong foundation in data engineering, production ML, feature pipelines, and model monitoring. I enjoy working on: • RAG pipelines (hybrid search, embedding strategies, chunking, re-ranking) • Agentic AI systems (LangGraph multi-agent workflow design) • LLM inference optimization (GPU batching, quantization, caching) • Retrieval quality evaluation and drift monitoring • Designing microservices for AI/ML in production • Fraud analytics, anomaly detection, scoring systems • Distributed data pipelines (Spark, Kafka, Databricks) If you’re working on LLM infrastructure, retrieval systems, agentic automation, or AI-driven decisioning, I’d love to connect and share ideas.
-
United States
Information Technology & Services
Microservices, Google Cloud Platform (GCP), MLOps, Fraud Detection, Model Building, Transformers, Prompt Engineering, Word Embeddings, Data Engineering, Apache Kafka, PySpark, Docker Products, Kubernetes, FastAPI, LangGraph, Retrieval-Augmented Generation (RAG), Large Language Models (LLM), LangChain, Predictive Maintenance, AWS SageMaker
Experience

AI/ML Engineer
Building LLM-powered internal assistants, multi-agent workflows, and RAG pipelines using LangChain, LangGraph, Pinecone, FAISS, and hybrid retrieval. Designed modular LLM architectures: embedding services, indexing pipelines, vector stores, ranking layers, context assembly, and GPU-backed inference. Delivered <400ms inference latency with FastAPI, async batching, quantization, and Redis caching. Improved document retrieval relevance by 32% Recall@10 and reduced operator workload across dispute/support workflows.

ML Engineer
Pune, Maharashtra, India
Developed ML models for fraud detection and anomaly scoring; improved AUC by ~12%. Built PySpark + Kafka data pipelines for high-volume transaction ingestion with 45% lower ETL latency. Created transformer-based NLP components for KYC/AML document analysis, reducing manual review by 40%. Contributed to real-time ML scoring systems, model versioning, drift monitoring, and automated retraining.
Education
Harshith G's Contact Information
Phone
Find the Right Leads
Find Verified Contact Data
What LeadContact does well
Find verified emails, phone numbers, and decision-makers with 98% accuracy.
Find Leads
Find the right people by company, role, industry, location, and more.
925M+ professional profiles

Find Emails
Access verified email addresses for your target contacts.
657M+ emails

Find Phone Numbers
Get cross-validated phone data from multiple top sources.
239M+ phone numbers

More Accurate. Lower Cost.
Find contact data in 1 tool with 98% accuracy
LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.
Great conversations start with the right contact.
It’s time to find yours.



