Summary
Overview
Work History
Education
Skills
Certification
Timeline
Generic

GIRI KAVITI

Jersey City,NJ

Summary

Highly skilled and passionate Data Engineer with over 12 years of experience in Cloud Computing, Big Data ecosystems, and data engineering. Expert in designing, developing, and optimizing data pipelines using Apache Spark, ETL technologies, and cloud platforms. Proven ability to transform massive datasets, manage distributed applications, and drive system performance improvements in enterprise environments. Demonstrated expertise across the entire SDLC, delivering robust data architecture and leading large-scale data migrations across banking and other sectors. Known for collaborating effectively with cross-functional teams and integrating cutting-edge technologies into real-world projects.

Overview

1
1
Certification
12
12
years of professional experience

Work History

Account Opening and Activation

JPMorgan Chase
Jersey City, NJ
07.2023 - Current
  • Developed monthly AOA dashboard unifying data from multiple lines of business, enhancing data-driven decision making and organizational alignment.
  • Designed and implemented dimension and fact tables for precise reporting and effective trend analysis.
  • Extracted and transformed data from diverse sources, staged it in AWS S3 using Java/Spark, and loaded it into Teradata and snowflake for advanced analytics.
  • Managed production infrastructure stability and optimization by resolving Performance issues in spark jobs , executing MQ certificate life cycle management , GSM/Vault migration and leading edge node to VSI platform migration
  • Participated in enterprise Disaster Recovery (DR) exercises by validating end-to-end data pipelines, restarting Spark jobs, verifying Kafka consumers, reconciling data across primary and secondary environments, and ensuring successful application recovery and production sign-off.

Online Data Mart

JPMorgan Chase
Jersey City, NJ
06.2021 - 06.2023
  • Orchestrated a large-scale Hadoop-to-AWS migration using S3, Spark, and Snowflake, significantly improving data processes and system performance.
  • Led the design and migration of ETL pipelines to ingest and process transactional and behavioral data from Cloudera to Hortonworks hadoop platform
  • Utilized SparkSQL within a proprietary Java/Spark framework, boosting data migration performance by 30%.

Sales Datahub

JPMorgan Chase
Jersey City, NJ
04.2020 - 05.2021
  • Developed end-to-end data extraction and transformation processes using Java-Spark and Talend, integrating data from Teradata, Oracle, and Salesforce to enhance data accessibility.
  • Developed end-to-end data extraction and transformation processes using Java-Spark and Talend, integrating data from Teradata, Oracle, and Salesforce.
  • Automated job scheduling with AutoSys, streamlining operational workflows and minimizing manual overhead.

CTRM DerivConnect

Zenith Services Inc.
Princeton, NJ
07.2019 - 03.2020
  • Gathered business requirements and identified technologies for CTRM-DerivConnect, enhancing design of a next-generation Big Data and analysis platform with Microservices.
  • Built data pipelines to ingest stock exchange and trading platform data into Hadoop/NoSQL environments, enabling data cleansing and analytics for accurate market predictions.
  • Designed systems to process and transform data into various formats (structured, semi-structured, JSON) for seamless consumption by SQL users, REST APIs, and data scientists.

Pivot to Hadoop

JPMorgan Chase
Hyderabad, India
07.2014 - 06.2019
  • Engineered reusable ETL framework with Ab Initio to streamline integration with Unified Data Services (UDS) and simplify development for internal teams.
  • Developed reusable components like CAIP READ/WRITE and CDC Type1/Type2, standardizing data ingestion and transformation processes across systems.
  • Developed reusable components like CAIP READ/WRITE and CDC Type1/Type2, standardizing data ingestion and transformation across systems.

Education

Bachelor of Engineering (BE) -

Osmania University
01-2014

Skills

  • Big Data Technologies: Apache Spark, Hadoop (HDFS, YARN, MapReduce), Hive, Kafka, Cassandra
  • Cloud Platforms: AWS (S3, Glue, Athena, EKS), Snowflake
  • Programming Languages: Java, Python, SQL, Shell Scripting (KSH)
  • Database Technologies: Teradata, DB2, Oracle, Cassandra, Salesforce
  • ETL Tools: Ab Initio, Talend
  • CI/CD & Automation: Jenkins, Jules, Control-M, AutoSys
  • Data Processing: SparkSQL, RDD, Java Spark, Datasets
  • Web Services: REST, SOAP APIs
  • NoSQL Databases: Cassandra, HBase

Certification

  • PCAP – Certified Associate Python Programmer
  • SnowPro Core Certification
  • AWS Certified Cloud Practitioner

Timeline

Account Opening and Activation

JPMorgan Chase
07.2023 - Current

Online Data Mart

JPMorgan Chase
06.2021 - 06.2023

Sales Datahub

JPMorgan Chase
04.2020 - 05.2021

CTRM DerivConnect

Zenith Services Inc.
07.2019 - 03.2020

Pivot to Hadoop

JPMorgan Chase
07.2014 - 06.2019

Bachelor of Engineering (BE) -

Osmania University