Summary
Overview
Work History
Education
Skills
Projects
Timeline
Generic

Mrunmay Muduli

Cincinnati

Summary

7 years of experience working with data across multiple domains. Over the years, built large-scale data pipelines, automated ETL processes and developed ML models using cloud platforms like AWS and Azure. One of my most impactful projects improved pipeline performance by 22% within six months, leading to faster and more reliable insights.

Overview

7
7
years of professional experience

Work History

Analytics Engineer

University of Cincinnati
Cincinnati
09.2025 - Current
  • Built production-ready Streamlit applications in Python to support analytics and business users, focusing on performance, caching, and clean UI patterns.
  • Designed reusable Streamlit components (filters, visualizations, exports) to standardize dashboard development.
  • Partnered with data science stakeholders to translate analytical requirements into scalable applications.

Senior Data Engineer

Capgemini
Bangalore
01.2025 - 08.2025
  • Migrated large-scale telecom analytics pipelines to Databricks and PySpark, enabling faster and more reliable analytical workloads.
  • Built and optimized ETL pipelines (100TB+) integrating Hive, MongoDB, Kafka, and S3, reducing processing delays by 22%.
  • Supported analytics and ML teams by curating reliable, reusable datasets for modeling and reporting use cases.

Application Developer

Pfizer
Remote
09.2024 - 12.2024
  • Migrated and transformed clinical trial data to CDISC SDTM/ADaM for regulatory submission within 4 months, accelerating compliance readiness about 40%.
  • Developed ARD datasets using R Programming and SQL, improving data traceability and reducing manual validation effort by 31%.

Senior Data Engineer

Persistent Systems
Bangalore
03.2024 - 08.2024
  • Built batch ETL workflows with optimized SQL, advanced Snowflake capabilities and data modeling to integrate data from 5+ sources within 2 quarters, driving process improvements and enhancing customer experience.
  • Improved data and content accuracy by 25% through customer specific logic and optimized delivery, boosting site responsiveness.

Data Scientist

Mu Sigma Business Solutions
Bangalore
01.2023 - 02.2024
  • Developed and maintained R Shiny dashboards for clinical trial and real-world evidence analytics, used by business and data science teams.
  • Automated analytics workflows using R, SQL, Power BI dashboards, and Redshift enable faster insights and improved experimentation cycles.
  • Built a Python POC for agent-based modeling and Bayesian networks, collaborating closely with advanced analytics teams.

Data Engineer

Great Software Laboratory
Pune
03.2021 - 01.2023
  • Developed an XGBoost model to predict customer churn with 81% accuracy, using feature-engineered telecom usage data.
  • Developed data pipelines from multiple sources using PySpark, AWS Glue, S3, Redshift, Step Functions, and web scraping to support a web app for 1,000+ users in Agile sprints.

Data Analyst

Tata Consultancy Services
Hyderabad
12.2018 - 03.2021
  • Developed PL/SQL database procedures and scripts to automate data processing for MANREGA, reducing runtime by 30%, and improving accuracy across 100K+ rural work records.
  • Built a SQL-based ETL pipeline to ingest and transform over 1 million daily financial transactions, integrating with reporting tools for fraud detection, and delivering timely customer insights.

Education

MS - Business Analytics

University of Cincinnati
Cincinnati, USA
08.2026

B.Tech - Mechanical Engineering

SOA University
Bhubaneswar, India
05.2018

Skills

  • Python
  • R Programming
  • Databricks
  • Kafka
  • PySpark
  • AWS
  • Azure
  • MongoDB
  • OLAP
  • OLTP
  • RDBMS (MySQL, PostgreSQL, Oracle)
  • Hive
  • Redshift
  • Snowflake
  • Airflow
  • Terraform
  • Dash
  • Power BI
  • R Shiny
  • Streamlit
  • Tableau
  • Gen AI
  • Machine Learning
  • Neural Network
  • Statistical modeling
  • CI/CD
  • Confluence
  • Docker
  • Github
  • Jira
  • Linux
  • Posit
  • Rest API
  • VS Code

Projects

Marketing analytics for a DTC brand to understand Media ROI, 08/2025, 12/2025, Executed geo-incrementality tests, MMM, and triangulation analysis to quantify media ROI, isolate seasonal effects, and forecast sales for a fast-growing DTC fashion brand. ML model for Anomaly detection, 11/2024, 12/2024, Detected anomalies using Random Forests and Decision Trees for a telecom client. Feedback Engine using NLP (Healthcare), 12/2023, 01/2024, Deployed Hugging Face Transformers to classify patient sentiments; integrated results into dashboards for clinical teams.

Timeline

Analytics Engineer

University of Cincinnati
09.2025 - Current

Senior Data Engineer

Capgemini
01.2025 - 08.2025

Application Developer

Pfizer
09.2024 - 12.2024

Senior Data Engineer

Persistent Systems
03.2024 - 08.2024

Data Scientist

Mu Sigma Business Solutions
01.2023 - 02.2024

Data Engineer

Great Software Laboratory
03.2021 - 01.2023

Data Analyst

Tata Consultancy Services
12.2018 - 03.2021

MS - Business Analytics

University of Cincinnati

B.Tech - Mechanical Engineering

SOA University
Mrunmay Muduli