Summary
Overview
Work History
Education
Skills
Projects & Achievements
Timeline
Generic

Reetesh Myneni

DeKalb,IL

Summary

Highly skilled Data Engineer with 10 years of experience in designing, building, and optimizing scalable data pipelines and architectures. Adept at working with big data technologies, cloud platforms, and ETL processes to drive data-driven decision-making. Experienced in collaborating with cross-functional teams to implement efficient data solutions. Strong expertise in Python, SQL, AWS.

Overview

12
12
years of professional experience

Work History

Data Engineer

Thermo Fisher Scientific Inc.
Boston, MA
01.2020 - 03.2025
  • Designed and developed robust ETL pipelines, handling terabytes of structured and unstructured data.
  • Optimized data storage solutions using AWS services such as S3, Redshift, and Glue, reducing data processing costs by 30%.
  • Implemented Apache Spark and Databricks for real-time data processing, improving analytics efficiency.
  • Worked closely with data scientists and analysts to enhance data accessibility and visualization tools.
  • Led a team of data engineers in developing scalable architectures to support business intelligence initiatives.

Data Engineer

Pfizer Inc.
New York, NY
05.2016 - 12.2019
  • Built and maintained data pipelines to integrate healthcare and pharmaceutical data from multiple sources.
  • Developed and optimized SQL queries and stored procedures for large-scale data processing.
  • Automated ETL workflows using Apache Airflow, reducing manual intervention by 40%.
  • Migrated on-premise data warehouses to AWS Redshift, improving query performance.
  • Ensured data governance and security compliance in accordance with industry standards.

Data Engineer

IBM
Raleigh, NC
07.2013 - 04.2016
  • Designed data models and implemented ETL solutions for enterprise applications.
  • Developed REST APIs for data ingestion and processing using Python and Flask.
  • Worked on Hadoop ecosystem tools such as Hive and HDFS for big data management.
  • Conducted performance tuning of SQL queries to enhance database efficiency.
  • Collaborated with software engineers to integrate data solutions into enterprise applications.

Education

Bachelor of Science - Information Technology

Prasad V Potluri Institute of Technolagy
Vijayawada
05-2012

Skills

  • Programming Languages: Python, SQL, Java, Scala
  • Big Data Technologies: Apache Spark, Hadoop, Hive
  • Cloud Platforms: AWS (Redshift, S3, Glue, Lambda), Azure, Google Cloud
  • ETL & Data Pipelines: Apache Airflow, Informatica, Talend
  • Databases: PostgreSQL, MySQL, MongoDB, Snowflake
  • DevOps & CI/CD: Docker, Kubernetes, Jenkins
  • Data visualization: Tablue, PowerBI

Projects & Achievements

  • Developed an end-to-end data pipeline for a predictive analytics model that improved forecasting accuracy by 25%.
  • Implemented a data lake architecture using AWS, reducing data duplication by 40%.
  • Led a migration project from legacy ETL tools to Apache Spark, enhancing processing speed by 5x.
  • Designed a real-time event processing system using Kafka and Spark Streaming, improving data ingestion by 70%.
  • Developed an automated data quality framework using Python, reducing data integrity issues by 35%.
  • Created a metadata-driven ETL framework that accelerated onboarding of new data sources by 50%.
  • Spearheaded the integration of machine learning pipelines with data engineering workflows, enabling predictive analytics capabilities.
  • Optimized an enterprise-wide data warehouse, improving query performance by 3x and reducing operational costs.

Timeline

Data Engineer

Thermo Fisher Scientific Inc.
01.2020 - 03.2025

Data Engineer

Pfizer Inc.
05.2016 - 12.2019

Data Engineer

IBM
07.2013 - 04.2016

Bachelor of Science - Information Technology

Prasad V Potluri Institute of Technolagy
Reetesh Myneni