

Data Engineer with 4+ years of experience designing, building, and supporting enterprise-scale data platforms on Microsoft Azure. Expertise in Azure Data Factory, Azure Databricks, PySpark, Apache Spark, SQL, and ADLS Gen2 to develop scalable ETL/ELT pipelines, cloud migrations, and high-performance data solutions. Experienced in production support, performance optimization, automation, and delivering reliable, AI-ready datasets for analytics, machine learning, and Generative AI. Strong background in data warehousing, Medallion architecture, data quality, governance, and cross-functional collaboration.
Languages: Python, Pyspark, SQL, Scala, Shell Scripting
Cloud: Microsoft Azure, Azure Data Factory (ADF), Azure Databricks, ADLS Gen2
Big Data: Apache Spark, PySpark, Delta Lake, ETL/ELT, Data Pipelines
Databases: SQL Server, Oracle, Teradata, MySQL, MongoDB
Concepts: Data Warehousing, Medallion Architecture (Bronze/Silver/Gold), Batch Processing, Data Migration, AI-ready Data Pipelines, Data Quality, CI/CD, Git, Azure DevOps, Data governance
Tools: Jira, SSMS, Oracle SQL Developer, Linux/Unix