Rida Fathima
Data Analytics Engineer @ Novanta Inc.
About
I am a dedicated enthusiast driven by a passion for solving real-world data problems using cutting-edge technologies, including cloud solutions and Machine Learning algorithms. I thrive on collaborating with cross-functional teams, encompassing Data Scientists, Data Engineers, and Business Analysts, to deeply understand requirements and deliver effective solutions. Fostering a culture of continuous learning and professional growth, and staying up to date with the industry trends, advancements, and emerging technologies. Knowledge and skills: - Proven expertise in data migration projects, with a focus on reliably migrating MS SQL data to GCP. - Experience in writing complex T-SQL queries, dynamic queries, indexes, constraints, views, functions and creating dynamic SSIS packages for data integration and ETL processes. - Proficiency in building Machine Learning Models, Time Series Models, Data Mining and Data Visualization. - Working knowledge of Apache Airflow and Apache Spark; Spark performance tuning - Experience in Cloud-based CI/CD process - Experience with Azure Databricks and Snowflake. - Well-versed with command line operations - Familiarity with AWS S3, Lambda, Redshift, Glue - Programming: SQL, Dynamic SQL, Python, R, PySpark, Pandas, NumPy, sklearn, C++, HTML - Visualization: Tableau, Power BI, Python (matplotlib, seaborn, plotly), Advanced Excel, Looker - Tools and Techniques: Jupyter, SSMS, SSIS, SSRS, ETL, Git, Visual Studio, RStudio - Database Technologies: MS SQL Server, MongoDB, MySQL, PostgreSQL - Machine Learning Techniques: Regression, Classification, Lasso and Ridge Regression, KNN, SVM, Decision Trees, Random Forest, Clustering, Anomaly Detection Models, PCA, Gradient Boost Algorithms, feature engineering, SMOTE, Data Mining, hyperparameter tuning, T-SNE plots, Spark ML, A/B Testing, Time Series Analysis, Resampling, Cross Validation, Reinforcement Learning, ETL, Sentiment Analysis, nltk, TF-IDF - Deep Learning Techniques: CNN, LSTM, RNN, keras, Pytorch, TensorFlow, BERT - Big Data Technologies: Spark, Hadoop, HDFS, MapReduce, Spark-SQL - Cloud Technologies: GCP, Azure, AWS, Databricks, Snowflake - Web Scraping: Selenium, Beautiful Soup
United States
Alpharetta
Information Technology & Services
PostgreSQL, Debezium, Data Build Tool (DBT), Git CI/CD, Kafka Streams, Informatica, MuleSoft Anypoint Platform, Shell Scripting, Unix, Data Warehousing, DataProc, Microsoft Power BI, Snowflake, Azure Data Factory, Azure Databricks, Azure Data Lake, Big Data, Data Engineering, Amazon Web Services (AWS), Apache Airflow
Experience

Data Analytics Engineer
Atlanta, GA
I work closely with engineering and finance teams in an Agile environment to align cloud data infrastructure with business goals. By centralizing transformations in Snowflake and optimizing Power BI integrations, I’ve improved data speed, reliability, and simplified access control. This role has strengthened my expertise in cloud-native architectures, data governance, and scalable data processing. -Handling enterprise-level data migration projects from legacy systems to Snowflake, establishing snowflake a production platform for Supply chain reports and maintaining data integrity. - Created Medallion architecture documentation to guide teams in adopting Snowflake and Power BI, streamlining data flow from ERP to reporting tables - immensely helping team during troubleshooting and understanding workflow. -Collaborated with data analysts, to implement dimensional schemas and deployed modular data warehouses on Snowflake lowering infrastructure costs by 18%. Also, streamlined access control eliminating the need for multi-source permissions. -Optimized Snowflake views and enabled direct Power BI connections, improving report refresh times, reducing transformation steps, simplifying troubleshooting, and boosting BI query performance by 42%. -Demonstrated ability to be flexible, adapt to changing project priorities, and manage a dynamic workload effectively. -Resolved daily tickets related to integrations, reporting, ETL, and data governance, while proactively coordinating with customers and assigning tasks based on requirements. -Investigated and resolved ETL issues during production incidents, ensuring minimal disruption and maintaining data pipeline reliability under time-sensitive conditions. -Created and managed user stories during sprint planning, ensuring timely completion within two-week sprints. Actively communicated with users and gathered business requirements required in performing the tasks.

Graduate Teaching Assistant
- Graduate teaching assistant for Data Management for Analytics course. - Grade the assignments of students and provide them feedback. - Assist the course instructor with insights on improving the course efficacy.

Data Engineer Intern
Atlanta, Georgia, United States
- Creating and implementing ETL pipelines using SSIS packages within the Visual Studio environment, ensuring seamless data transformation and integration. - Creating and executing Stored Procedures in Google Big Query to efficiently load SQL Server data into the cloud, enabling fast and effective data management and retrieval. - Migrated existing pipelines and processes to the cloud (Google Cloud Platform) - Employed Cron Job schedulers to automate repetitive tasks related to table creation, loading data, and populating tables in GCP. - Developing an automation system to calculate the row count of tables in SQL and GCP, and configured it to send email notifications via SQL whenever a discrepancy in row count is detected. - Worked on Tableau for reporting data quality check analysis, created process flow diagrams and maintained detailed documentation of transformations and data lineage. - Actively participating in daily stand-up calls and demonstrating adherence to Agile Framework by engaging in sprint planning, retrospective and review sessions.

Data Science Research Assistant
Atlanta, Georgia, United States
- Created an ETL Pipeline using Apache Airflow and PySpark in GCP to detect the anomalous data points in Mercedes Benz’s SAP Data Warehousing Systems. - Implemented advanced Machine Learning Models (LOF, IForest) and utilized T-SNE Plots to visualize high-dimensional data and enable the identification of Local outliers within the dataset. - Developed an interactive Dashboard on the cloud using Dash API, empowering Data Specialists to seamlessly identify and investigate anomalies in the data. - Produced reports for SQL Administrators to ensure the data accuracy and facilitate a robust product development process - Worked in a team and demonstrated Agile principles during Sprint Reviews, showcasing completed work and gathering feedback for continuous improvement.

Data Analytics Engineer Intern
FRR IT Services
Hyderabad
- Developed an ETL pipeline to Automate the ingestion of flat files into Microsoft SQL Server, resulting in significant time savings (from 100 records per hour to 2000 records per hour). - Utilized data cleansing techniques to accurately identify attributes and tuples within text data, resulting in improved data integrity. - Implemented dynamic SQL statements to create databases, schemas, and tables to streamline the database creation process - Worked on recursive stored procedures, materialized views, functions and performance tuning.
Education
Rida Fathima's Contact Information
Phone
Find the Right Leads
Find Verified Contact Data
What LeadContact does well
Find verified emails, phone numbers, and decision-makers with 98% accuracy.
Find Leads
Find the right people by company, role, industry, location, and more.
925M+ professional profiles

Find Emails
Access verified email addresses for your target contacts.
657M+ emails

Find Phone Numbers
Get cross-validated phone data from multiple top sources.
239M+ phone numbers

More Accurate. Lower Cost.
Find contact data in 1 tool with 98% accuracy
LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.
Great conversations start with the right contact.
It’s time to find yours.



