Summary
Overview
Work History
Education
Skills
Affiliations
Timeline
Generic

Pavan Kalyan Chippalapalli

Milwaukee,WI

Summary

Results-driven Data Engineer with 5 years of experience in designing and maintaining scalable data pipelines and cloud solutions. Expertise in ETL/ELT processes, data warehousing, and big data technologies, particularly within Microsoft Azure. Proven skills in Python, SQL, and data analytics tools, enhancing data platforms and optimizing workflows for Healthcare and Banking sectors.

Overview

7
7
years of professional experience

Work History

Data Engineer

McKesson Corporation
Irving, Texas
03.2025 - Current
  • Developed PySpark notebooks to efficiently process structured and semi-structured healthcare data, enhancing data accessibility for analytics.
  • Implemented Delta Lake architecture to improve data reliability and performance.
  • Designed and developed scalable ETL pipelines using Azure Data Factory and Azure Databricks.
  • Built data ingestion pipelines from SQL Server, APIs, and flat files into Azure Data Lake Storage Gen2.
  • Optimized Spark jobs and SQL queries, reducing ETL execution time by more than 35%.
  • Integrated Snowflake with Azure services, streamlining enterprise reporting and analytics for improved decision-making.
  • Created SQL stored procedures, views, and performance-tuned queries to optimize database interactions and enhance query performance.
  • Implemented CI/CD pipelines using Azure DevOps.
  • Participated in Agile ceremonies including Sprint Planning, Daily Stand-ups, Sprint Reviews, and Retrospectives.

Environment: Azure Data Factory, Azure Databricks, Azure Synapse Analytics, Azure Data Lake Storage Gen2, Python, SQL, PySpark, Snowflake, Azure DevOps, Git, Power BI.

Data Engineer

Andhra Bank
CHIRALA, ANDHRA PRADESH
01.2020 - 08.2023
  • Built PySpark applications for processing large-scale financial datasets.
  • Developed scalable data ingestion pipelines from multiple enterprise systems.
  • Designed and developed enterprise ETL pipelines using Azure Data Factory.
  • Optimized SQL queries for high-performance reporting.
  • Implemented incremental loading strategies using watermarking techniques.
  • Designed dimensional data models to enhance business intelligence solutions.
  • Built automated data validation and reconciliation frameworks.
  • Collaborated with cross-functional teams to deliver enterprise data solutions.
  • Ensured data governance, security, and quality standards across the platform.

Environment: Azure Data Factory, Azure Databricks, SQL Server, Snowflake, Azure Synapse Analytics, Azure DevOps, Python, PySpark, Power BI, Git.

Education

Master's Degree - Information Technology

University of Wisconsin, Milwaukee
Milwaukee, WI
12-2024

Bachelor of Engineering - Electronics And Communication Engineering

Sathayabama University
Chennai, India
06-2020

Skills

  • Azure Data Factory (ADF)
  • Azure Databricks
  • Azure Synapse Analytics
  • Azure Data Lake Storage Gen2
  • Snowflake
  • ETL / ELT
  • PySpark
  • Apache Spark
  • Delta Lake
  • SQL
  • Python
  • Azure DevOps
  • CI/CD
  • Git

Affiliations

  • Enterprise Healthcare Data Platform
    * Designed cloud-based ETL pipelines processing more than 10 million healthcare records.

    * Improved data quality using automated validation and monitoring.

    *Reduced ETL processing time by approximately 40%.

    *Built scalable Azure Data Lake architecture supporting enterprise analytics.

Timeline

Data Engineer

McKesson Corporation
03.2025 - Current

Data Engineer

Andhra Bank
01.2020 - 08.2023

Master's Degree - Information Technology

University of Wisconsin, Milwaukee

Bachelor of Engineering - Electronics And Communication Engineering

Sathayabama University
Pavan Kalyan Chippalapalli