Summary
Overview
Work History
Education
Skills
Certification
Volunteer work
AWARDS
Timeline
Generic

SAIKRISHNA BONALA

Scottsdale,AZ

Summary

Strategic Senior Data Engineer and Team lead with 13+ years of experience developing scalable, cloud-native data ecosystems. Led implementation of Generative AI tools like Microsoft Copilot and ChatGPT to enhance delivery speed and operational efficiency. Drove DataOps automation and managed enterprise roadmaps, optimizing costs and governance to empower cross-functional data teams.

Overview

1
1
Certification
13
13
years of professional experience

Work History

Sr. Data Engineer & Team Lead

Marsh & McLennan
11.2021 - Current
  • Architected an end-to-end ingestion pipeline for Cyber Self-Assessment data using Databricks Lakehouse Platform (Unity Catalog, Data bricks SQL, Spark SQL, Python) to ingest raw data from mongoDB/PostgreSQL into a Medallion Architecture, utilizing Data bricks Autoloader for incremental file processing, Delta Live Tables (DLT) for automated transformation, and Data bricks Workflows with Git Repositories for CI/CD orchestration, reducing ingestion latency by 60%.
  • Led a team of Data engineers to successfully ingest and integrate Cyber Self-Assessment insurance data into core systems, ensuring 100% data accuracy.
  • Developed Spark pipelines for cyber insurance data processing and implemented Delta Lake for effective data management and performance enhancement.
  • Designed and implemented serverless data pipelines using AWS Lambda, AWS Step Functions and AWS Glue to orchestrate complex workflows and process terabytes of data daily.
  • Monitored and optimized Databricks cluster performance by tracking metrics to identify bottlenecks, improving resource utilization.
  • Partnered directly with Product Owners to translate complex business requirements into actionable technical specifications, driving successful product delivery.
  • Used Microsoft Copilot, ChatGPT, and other AI-assisted tools in day-to-day data engineering work to accelerate development across Databricks, PySpark, Python, SQL, PostgreSQL, MongoDB, and AWS.
  • Leveraged AI for code generation, debugging, query optimization, documentation, and root-cause analysis, improving engineering productivity and reducing turnaround time for development tasks.
  • Applied AI-assisted support in routine activities such as schema mapping, data validation, pipeline troubleshooting, transformation logic development, and knowledge discovery across enterprise data workflows.
  • Utilized Databricks AI/BI capabilities to enhance business-user access to insights, enabling faster analysis through natural-language-driven exploration.

Big Data Engineer

TATA Consultancy Services, India
10.2015 - 11.2021
  • Designed and implemented serverless data pipelines using AWS Lambda, AWS Step Functions and AWS Glue to orchestrate complex workflows and process terabytes of data daily.
  • Developed Spark applications using Python on Azure Databricks to enhance data interaction with MySQL and facilitate access to Hive tables through Spark SQL and Hive contexts.
  • Loaded batch data into Delta Lake, enabling reliability, performance, and interoperability across the Azure analytics engines.
  • Designed and created Hive external tables using shared meta-store with static partitioning, dynamic partitioning, and bucketing.
  • Created Spark scripts using Python shell commands to meet specified project requirements and ensure alignment with project goals.
  • Implemented Agile methodology to streamline project delivery, fostering collaboration between cross-functional teams and ensuring efficient development cycles.
  • Analyzed customer DNA mapping document from Business Analyst and clarified logic with DM team to ensure comprehensive understanding of dataflow.
  • Environment: Hadoop, Hive, Spark SQL, HDFS, PySpark, Azure Data Factory, Azure Data Lake Storage, Azure Synapse Analytics, Azure Databricks, Delta Lake, Azure SQL Database, GitHub, Azure DevOps, AWS EMR, AWS S3, AWS Anthena, AWS Lambda, AWS Step and AWS GLUE, Amazon Redshift, Apache Flink, MongoDB.

Data Analyst

Equifax, India
04.2013 - 10.2015
  • Applied data visualization techniques and designed interactive dashboards using SAS Visual Analytics to present complex reports, charts, summaries, and graphs to team members and stakeholders..
  • Developed modules to analyze trends in credit card usage, focusing on key factors such as spends, balance, acquisition, and validation checks, enhancing understanding of customer behavior.
  • Analyze the trend of credit card customers on various aspects like CIF (Card in force), New cards opened, Spends, ATS, Outstanding.
  • Implemented ETL processes and optimized SQL queries to streamline data extraction and merging from SQL server database, improving data accessibility for analysis.
  • Implemented ETL process wrote and optimized SQL queries to perform data extraction and merging from SQL server database.
  • Designed and created Hive external tables using shared meta-store with static and dynamic partitioning and bucketing, facilitating efficient data management and querying.
  • Worked with various file formats, including CSV, TXT, and fixed width, to efficiently load data from multiple sources into raw tables.
  • Managed data from multiple sources and performed HDFS maintenance.
  • Environment: Hadoop, Hadoop Distributed File System, MapReduce, Hive, Spark, Sqoop, Power BI, Shell Script, SQL, AutoSys, BASE SAS, SAS MACROS, SAS Enterprise guide.

Education

Bachelor 's - Electronics and Communications

JNTUH India
India

Skills

  • Generative AI & Productivity: Microsoft Copilot, GitHub Copilot, ChatGPT, AI-assisted coding, prompt engineering, debugging, documentation automation, SQL optimization, pipeline troubleshooting
  • Data Ecosystems: Databricks, AWS cloud service expertise, Azure Data Services
  • Databricks stack: MLflow, Delta Lake, Unity Catalog
  • DataOps & Engineering: Automated testing,code quality checks, reusable deployment templates
  • Big Data Technologies: HDFS, MapReduce, Hive, Spark streaming, Spark, Sqoop
  • Data pipeline orchestration
  • Languages: SQL, PL/SQL, Python, HiveQL, PySpark, SAS
  • Databases: MS SQL Server, MS Excel, MS Access, Oracle 11g/12c
  • Experienced with NoSQL database management
  • Visualization tools: Tableau, SAS
  • Data Governance: Metadata management, data lineage, data stewardship, access controls

Certification

1.AWS Certified Data Engineer - Associate

2.SAS Certified Professional: Advanced Programming Using SAS 9.4

Volunteer work

  • Coordinated community outreach programs to increase volunteer engagement and participation.
  • Dedicated volunteer providing logistical support and food distribution for Phoenix-based non-profits, including St. Mary's Food Bank and Andre House.

AWARDS

1. Customer champion Award

2. Empathy in Action Awards

3. PAT on the Back Awards

4. Star Performer Award

Timeline

Sr. Data Engineer & Team Lead

Marsh & McLennan
11.2021 - Current

Big Data Engineer

TATA Consultancy Services, India
10.2015 - 11.2021

Data Analyst

Equifax, India
04.2013 - 10.2015

Bachelor 's - Electronics and Communications

JNTUH India
SAIKRISHNA BONALA