Mahalaxmi Anumula

Mahalaxmi Anumula

Data Engineer @ Pfizer

About

Results-driven Data Engineer with ~5 years of experience building scalable data integration, transformation, and analytics solutions across AWS and Azure. I specialize in designing high-performance ETL/ELT pipelines using PySpark, AWS Glue, and Azure Data Factory, and architecting modern data lakehouse systems enabling fast, reliable analytics. Proficient in Python, SQL, Apache Spark, and experienced with Snowflake, Redshift, Tableau, and Power BI, I combine strong engineering skills with analytical thinking to deliver high-impact data solutions. I’ve deployed machine learning models, implemented workflow orchestration with Airflow, automated CI/CD processes, and ensured high-quality, well-governed data environments. I thrive in Agile teams, collaborate effectively with cross-functional partners, and enjoy solving complex data challenges that translate directly into business value.

Country

United States

City

Bellevue

Industry

Computer Software

Skill

Continuous Integration and Continuous Delivery (CI/CD), Kubernetes, Terraform, Version Control, Gitlab, Amazon Redshift, Amazon Elastic MapReduce (EMR), Amazon CloudWatch, dynamodb, PostgreSQL, Snowflake, Apache Spark, Apache Spark Streaming, hdfs, MongoDB, Jenkins, Git, Amazon EC2, Amazon S3, AWS Lambda

Experience

Pfizer

Data Engineer

Pfizer

LinkedIn
2023-9 - Present · 3 yrs 1 mo

United States

• Designed and developed scalable ETL pipelines using AWS Glue and PySpark to process healthcare claims and patient data, ensuring HIPAA compliance and reducing pipeline runtime by 30%. • Architected a data lakehouse solution on AWS (S3, EMR, Athena, Redshift) for storing, transforming, and querying large-scale clinical and operational data. • Implemented data workflow orchestration with Apache Airflow, automating ingestion and transformation jobs across multiple data domains. • Developed predictive machine learning models using Python, Scikit-learn, and TensorFlow to identify high-risk patients and optimize care pathways, improving prediction accuracy by 20%. • Created automated Tableau dashboards to visualize claims performance metrics and clinical KPIs, enabling executive teams to track real-time healthcare trends. • Integrated data validation and monitoring frameworks using Datadog and SQL, ensuring continuous quality and consistency of production data. • Collaborated within Agile Scrum teams using Jira and Confluence to manage sprints, document pipelines, and deliver analytics solutions in production environments. • Established CI/CD pipelines with Jenkins and Git, improving deployment efficiency and version control across ETL environments.

Mindtree

Data Engineer

Mindtree

LinkedIn
2020-1 - 2022-7 · 2 yrs 7 mos

India

• Engineered Azure-based data pipelines using Data Factory and Databricks to integrate and transform large volumes of enterprise data from on-premise SQL Server systems. • Designed and implemented data models and schemas in Azure Synapse Analytics, optimizing query performance for downstream analytics and reporting. • Developed robust ETL workflows using Informatica and SSIS, automating end-to-end data movement and reducing manual intervention by 50%. • Performed advanced data wrangling and preprocessing using Python (Pandas, NumPy) and SQL for high-quality, analytics-ready datasets. • Leveraged Apache Spark for distributed data processing, achieving a 40% improvement in transformation speed for large transactional datasets. • Built interactive Power BI dashboards for procurement and supply chain analytics, driving cost optimization and enhancing supplier performance visibility. • Practiced Agile SDLC methodology, contributing to sprint reviews, maintaining documentation in Jira, and coordinating across global development teams.

Education

DePaul University

DePaul University

LinkedIn

Data Science

Osmania University, Hyderabad

Osmania University, Hyderabad

LinkedIn

Information Technology

Mahalaxmi Anumula's Contact Information

Email

******@***.com

Phone

(**) *** ****

Find the Right Leads
Find Verified Contact Data

Try with: Jensen Huang @ nvidia.com Click to autofill
LeadContact awards, five-star ratings, and GDPR compliance badges

What LeadContact does well

Find verified emails, phone numbers, and decision-makers with 98% accuracy.

Find Leads

Find Leads

Find the right people by company, role, industry, location, and more.

925M+ professional profiles

Find Leads
Find Emails

Find Emails

Access verified email addresses for your target contacts.

657M+ emails

Find Emails
Find Phone Numbers

Find Phone Numbers

Get cross-validated phone data from multiple top sources.

239M+ phone numbers

Find Phone Numbers

More Accurate. Lower Cost.

Find contact data in 1 tool with 98% accuracy

LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.

LeadContact Logo
Competitor Tools

All these = $289 per month

Great conversations start with the right contact.

It’s time to find yours.