Ramya S Siddela
Data Engineer @ VENNISON TECHNOLOGIES INC
About
I'm a Data Engineer with 5+ years building cloud-scale Lakehouse architectures across healthcare, retail, and banking the kinds of systems where bad data or slow pipelines have real consequences. My stack is Snowflake, Databricks, Azure, and dbt, but the actual job is making data trustworthy, fast, and usable for the people who depend on it. 𝗪𝗵𝗮𝘁 𝗜'𝘃𝗲 𝗱𝗲𝗹𝗶𝘃𝗲𝗿𝗲𝗱: → Reduced full-refresh compute by ~60% using Snowflake Streams & Tasks CDC → Improved query performance 30–70% via micro-partition pruning and clustering → Cut JSON/VARIANT query execution from 45s → 12s (73% faster) → Processed 200M+ records/day across multi-domain Lakehouse pipelines → Built a GenAI analytics app on Snowflake Cortex — slashed manual review work by 80% → Architected Bronze/Silver/Gold Medallion pipelines on Delta Lake at production scale 𝗛𝗼𝘄 𝗜 𝘄𝗼𝗿𝗸: I'm Snowflake-first and cost-obsessed. I design multi-account environments (DEV/QA/PROD), implement proper RBAC and governance for HIPAA-sensitive data, and build CI/CD pipelines so deployments are boring which is exactly what you want in production.𝗖𝗲𝗿𝘁𝗶𝗳𝗶𝗲𝗱: SnowPro Core (COF-C02) · Databricks Data Engineer Associate · AWS Cloud Practitioner · Gen AI with Snowflake 𝗥𝗶𝗴𝗵𝘁 𝗻𝗼𝘄: Based in Dallas, TX. On STEM OPT — authorized to work in the US for up to 3 years Open to: Senior Data Engineer · Analytics Engineer · Cloud Data Engineer Available for: Full-time or C2C · Remote, hybrid, or on-site If you're hiring or know someone who is let's talk: ramya.siddela09@gmail.com
United States
Dallas-Fort Worth Metroplex
Computer Software
ETL/ELT, Medallion Architecture, CI/CD, Data Build Tool (DBT), dbt core, Snowflake, Apache Spark, Microsoft Azure, Databricks Jobs / Workflows, Snowflake SQL, Data Engineering, Natural Language Processing (NLP), Azure Databricks, PySpark, Delta Lake (MERGE, OPTIMIZE, ZORDER), Unity Catalog, Cloud Computing, Agile methodology, Technical Research, cloud security across network, application, and data layers
Experience

Data Engineer
McKinney, TX
• Architected Azure–Databricks–Snowflake Lakehouse ingesting 200M+ records/day across sales, inventory, and banking domains reduced data latency from 4+ hours to under 10 minutes using Snowpipe + ADF • Implemented Snowflake Streams & Tasks CDC architecture, eliminating full-refresh workloads and cutting pipeline compute by ~60%; automated orchestration reduced manual intervention to zero • Optimized Snowflake warehouse performance through clustering, micro-partition pruning, and warehouse auto-suspend — delivered 30–35% faster queries and 12–18% reduction in monthly compute spend • Designed and implemented Bronze/Silver/Gold Medallion architecture on Delta Lake + PySpark, applying OPTIMIZE and ZORDER to high-scan tables cut average query scan time by 40% and eliminated downstream pipeline bottlenecks across 3 reporting domains. • Developed dbt Core transformation layer (staging → intermediate → mart) with incremental models, automated data quality tests, and DAG dependency management — reduced pipeline failures by ~30% • Designed and deployed CI/CD pipelines (Azure DevOps + Git) for Snowflake schema changes, dbt runs, and Databricks job promotions across DEV/QA/PROD environments • Implemented RBAC, column-level masking policies, and row-level security across Snowflake multi-account environment (DEV/QA/PROD) zero data governance incidents across healthcare and banking workloads.

Snowflake Engineer
Milwaukee, WI
• Built HIPAA-compliant Snowflake ELT pipelines for healthcare claims, provider, and eligibility data (40K+ records/day); implemented Snowpipe auto-ingest, cutting refresh latency from 3+ hours to under 5 minutes • Rewrote VARIANT/JSON parsing queries — reduced execution time from ~45 seconds to ~12 seconds (73% improvement), directly improving dashboard load time for clinical stakeholders • Enforced HIPAA-compliant data governance across all PHI tables implemented RBAC, column-level masking, and secure views, achieving zero data governance violations across 40K+ daily healthcare claims records.

Junior Snowflake Developer
Hyderabad
• Contributed to SQL Server → Snowflake enterprise migration for customer, account, and transaction datasets and improved analytics data availability by ~30% and reduced ad-hoc query turnaround from days to under 2 hours. • Built external stage integrations with AWS S3 for automated daily ingestion via Snowpipe; eliminated manual file-load process that previously required 2+ hours of engineer time per day • Built fraud analytics SQL views and stored procedures used daily by the risk team — identified clustering optimizations that reduced scan volume on 500M+ row transaction tables by ~45%, directly improving report generation time. • Automated daily data ingestion via AWS S3 external stages and Snowpipe, eliminating 2+ hours of manual file-load engineering time per day.
Ramya S Siddela's Contact Information
Phone
Find the Right Leads
Find Verified Contact Data
What LeadContact does well
Find verified emails, phone numbers, and decision-makers with 98% accuracy.
Find Leads
Find the right people by company, role, industry, location, and more.
925M+ professional profiles

Find Emails
Access verified email addresses for your target contacts.
657M+ emails

Find Phone Numbers
Get cross-validated phone data from multiple top sources.
239M+ phone numbers

More Accurate. Lower Cost.
Find contact data in 1 tool with 98% accuracy
LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.
Great conversations start with the right contact.
It’s time to find yours.


