Summary
Overview
Work History
Education
Skills
Websites
Timeline
Generic

Pavan Krishna VUNNAM

Bridgeport,CT

Summary

AI Data Engineer with 3+ years of experience designing, implementing, and optimizing enterprise-scale data processing systems and AI-powered cloud architectures across AWS and Azure ecosystems. Proven expertise in building scalable ETL and ELT pipelines using AWS Glue, Lambda, Step Functions, Azure Data Factory, and Databricks. Specialized in Retrieval-Augmented Generation (RAG), semantic search architectures, embeddings pipelines, and LLM integration using Amazon Bedrock, Kendra, and OpenSearch. Experienced in architecting AI-ready data lakes with partitioning, encryption, lifecycle management, and metadata governance. Strong background in distributed data processing using PySpark for large-scale transformations and feature engineering. Hands-on experience implementing Attribute-Based Access Control (ABAC) using IAM policies and S3 object tagging for ITAR and EAR compliance. Proficient in Python, SQL, and API integrations for automation, transformation, and data orchestration. Experienced in processing large volumes of structured and unstructured files including PDF, CSV, DAT, TXT, and image-based documents. Skilled in cross-cloud architecture integrating AWS and Azure platforms for unified analytics and machine learning workflows. Strong knowledge of observability, logging, audit trails, and monitoring using CloudWatch and structured tagging frameworks. Proven ability to reduce operational overhead through automation, performance optimization, and dashboard engineering. Passionate about building secure, scalable, AI-driven data ecosystems that align engineering excellence with business impact.

Overview

4
4
years of professional experience

Work History

Data Engineer

Aionix11
New Haven, CT
11.2024 - Current
  • Designed and deployed serverless ETL pipelines using AWS Glue, Lambda, and Step Functions for structured and unstructured data ingestion.
  • Built enterprise data lakes in Amazon S3 and Azure ADLS Gen2 with partitioning, encryption, lifecycle policies, and metadata tagging.
  • Developed distributed data transformation workflows using Databricks and PySpark for large-scale analytics and feature engineering.
  • Created automated ingestion workflows from REST APIs and flat files into Redshift, Snowflake, and Synapse.
  • Implemented Attribute-Based Access Control (ABAC) using IAM principal tags and S3 object tagging for secure data governance.
  • Developed AI Knowledge Search Platform integrating Amazon Kendra, Bedrock, and OpenSearch for semantic search and contextual Q&A.
  • Automated metadata tagging, logging, and audit trails using Python and CloudWatch for observability and compliance.
  • Engineered optimized data models supporting executive dashboards in Power BI and QuickSight, reducing reporting effort by 70 percent.
  • Implemented CI/CD pipelines using Terraform and GitHub Actions for infrastructure automation.
  • Architected cross-cloud data pipelines connecting AWS, Databricks, and Azure Synapse for unified analytics.

Statistics Tutor

University of Bridgeport
Bridgeport, CT
09.2023 - 06.2024
  • Delivered graduate-level instruction in regression, probability, and data analytics using Python and R.
  • Designed reproducible Jupyter Notebook workflows for statistical modeling and automation.
  • Guided students in data cleaning, feature engineering, and predictive modeling.

Junior Data Analyst

Compsoft Technologies
Bangalore, India
03.2022 - 07.2023
  • Designed ETL pipelines ingesting customer and sales data into centralized MySQL and Power BI systems.
  • Automated data transformations and reporting using Python and SQL, reducing manual tasks by 40 percent.
  • Developed regression and time-series forecasting models improving business prediction accuracy by 18 percent.
  • Built executive dashboards integrating SQL queries and KPI metrics.
  • Implemented validation scripts and reconciliation procedures to ensure data integrity.

Education

Master of Science - Computer Science

University of Bridgeport
Bridgeport, CT
01-2025

Bachelor of Engineering - Electronics and Communication

Atria Institute of Technology
Bangalore, India
01-2022

Skills

  • Programming: Python (Pandas, NumPy, PySpark, boto3, Scikit-learn), SQL, R, Shell
  • Cloud Platforms: AWS (S3, Glue, Lambda, Step Functions, Redshift, Athena, CloudWatch, IAM, Bedrock, Kendra, OpenSearch, EMR) Azure (Data Factory, Databricks, Synapse Analytics, ADLS Gen2)
  • Data Engineering & Orchestration: Terraform, Airflow (MWAA), dbt, EventBridge, Snowflake
  • Databases: Redshift, Snowflake, MySQL, PostgreSQL, Synapse Analytics, DynamoDB
  • Visualization: Power BI, Tableau, Amazon QuickSight
  • AI & Machine Learning: Retrieval-Augmented Generation (RAG), Embeddings Pipelines, NLP, Regression, Forecasting, Bedrock Agents, MLflow (basic)
  • DevOps & Collaboration: GitHub, CI/CD, JIRA, Agile Methodology

Timeline

Data Engineer

Aionix11
11.2024 - Current

Statistics Tutor

University of Bridgeport
09.2023 - 06.2024

Junior Data Analyst

Compsoft Technologies
03.2022 - 07.2023

Master of Science - Computer Science

University of Bridgeport

Bachelor of Engineering - Electronics and Communication

Atria Institute of Technology