Summary
Overview
Work History
Education
Skills
Websites
Certification
Immigration Status
Timeline
Generic

Monish Joshy

Missouri City

Summary

Dynamic Senior Data Analyst with extensive experience specializing in data integration and transformation using Azure Synapse and Data Factory. Proven track record of optimizing ETL processes, enhancing data quality, and driving significant improvements in reporting accuracy. Strong analytical skills combined with effective team leadership capabilities.

Overview

12
12
years of professional experience
1
1
Certification

Work History

Lead Data Analyst and Engineer

Infinite/Scan Health
10.2025 - Current
  • Created complex T-SQL queries and stored procedures for large datasets.
  • Managed project team supporting Scan healthcare in development and support of EDW team.
  • Designed, architected, developed, and supported intricate SSIS ETL applications of Fusion BPA.

Senior Data Analyst & Data Engineer

American Family Insurance through Cloudpaths/EY
Houston
01.2025 - 09.2025
  • Developing and supporting data pipelines, ETL processes, and data warehousing solutions in Azure and on-premises environments.
  • Developed SQL, PLSQL, and Python scripts for insurance document generation.
  • Supported Informatica ETL pipelines for multiple lines of business.
  • Designed complex T-SQL queries and stored procedures for large datasets.
  • Created data pipelines in Databricks for data transformation and loading.
  • Worked closely with business and technical teams to define integration requirements and deliver scalable solutions.
  • Extract Transform and Load data from Sources Systems to Azure Data Storage services using a combination of Azure Data Factory, T-SQL, Spark SQL and U-SQL Azure Data Lake Analytics. Data Ingestion to one or more Azure Services - (Azure Data Lake, Azure Storage, Azure SQL, Azure DW) and processing the data in In Azure Databricks.
  • Data Ingestion to Azure Services - Azure Data Lake, Azure Storage, Azure SQL, Azure DW, and processing the data in Azure Databricks.
  • Debugging Java front end application code.

Senior Data Analyst / Data Engineer

Amsive
Remote
05.2022 - 11.2024
  • Designed and maintained Azure Synapse pipelines for diverse data sources.
  • Built and optimized data pipelines and reporting solutions using Azure Synapse, T-SQL, and SSIS, ensuring high data quality and performance.
  • Built data pipelines and reporting solutions using pure T-SQL, handling data aggregation, transformation, and presentation from scratch without relying on third-party tools.
  • Automated Azure Pipelines using trigger and monitored the progress of pipeline execution, also deployed SSIS within Integration Catalog and scheduled the SSIS Packages to run at certain time using SQL Jobs.
  • Developed SSIS packages that utilized custom T-SQL scripts for data extraction, transformation, and loading (ETL) to ensure seamless integration between disparate data sources and target systems.
  • Wrote T-SQL queries and stored procedures within SSIS Data Flow tasks to perform complex data transformations, aggregations, and lookups on large datasets.
  • Managed and processed large data sets across multiple projects, ensuring data integrity and reliability through efficient SQL and ETL operations.
  • Created resilient pipeline monitoring using Python scripts that parse logs from Databricks Spark jobs, automatically restarting failed tasks and sending alerts to key stakeholders, increasing system uptime by 30%.
  • Optimized and redesigned database schemas, views, and objects post-migration to Azure Synapse Analytics, leveraging partitioning, indexing, and columnar storage for performance improvements.
  • Migrated complex T-SQL queries, stored procedures, and views from on-prem SQL Server to Azure Synapse, ensuring compatibility and improved performance using Synapse’s distributed architecture.
  • Optimized data transformations within MS Synapse / Azure Synapse and Spark SQL to handle complex, high-volume data sets for healthcare and finance clients, reducing processing times.
  • Processed large (Billions) data sets with T-SQL queries to support business intelligence and data warehousing initiatives, enhancing reporting accuracy for healthcare and finance domains.
  • Architected and deployed MS Synapse / Azure Synapse Pipelines to support large-scale data integration from various sources, driving real-time data availability for analytics.
  • Developed azure pipelines Experience with unity Catalog and performing complex transformations within data bricks.
  • Optimized and refactored existing T-SQL code, improving query execution time by 40% through indexing strategies, query rewriting, and the use of advanced T-SQL technique.
  • Designed and implemented scalable ETL pipelines using Apache Spark on Databricks, optimizing data processing workflows across cloud environments.
  • Integrated Delta Lake for structured streaming and batch processing to ensure ACID compliance and real-time analytics capabilities.
  • Automated data ingestion and transformation processes using Databricks Notebooks and job scheduling, significantly reducing manual intervention and improving pipeline reliability.
  • Leveraged Databricks SQL and Unity Catalog to enable secure, governed access to datasets and support business intelligence reporting.
  • Developed Pyspark notebooks to execute marketing campaigns of different healthcare clients Used ETL from different source systems to Azure storage services using a combination of azure pipelines, T-SQL, PostgreSQL, Spark SQL.
  • Implemented CI/CD pipelines using Azure DevOps to manage the deployment of MS Synapse / Azure Synapse and ADF pipelines, ensuring seamless, reliable updates to data workflows.
  • Developed robust ETL workflows to process and transform data from multiple sources, ensuring clean and structured data for downstream analytics.
  • Optimized SQL code, reducing query times by 40%.
  • Automated pipeline deployment and monitoring with Azure DevOps.
  • Migrated on-premises SQL Server to Azure Synapse, enhancing performance.
  • Developed Spark and Databricks workflows for large-scale data processing.
  • (Clients: BCBS/ HAP/ Wellmark/ Molina / Scott White)

Data Analyst / Database Developer

Texas Education Agency
Austin
03.2020 - 05.2022
  • Engineered ETL processes and data validation routines for state-wide education data systems.
  • Developed complex stored procedures in PLSQL.
  • Managed database backup, recovery, and tuning for large datasets.
  • Enhanced data quality and reporting accuracy through advanced SQL techniques.

Technology Lead – US/Data Analyst

Infosys
Houston
03.2019 - 03.2020
  • Led development of ETL packages, Power BI dashboards, and data validation tools for financial and operational data.
  • Created Power BI reports with advanced DAX and visualization techniques.
  • Automated data quality checks and implemented security frameworks with Azure Key Vault.
  • Led migration and optimization of data workflows to Azure cloud environments.
  • Involved in the development of ETL Packages using SSIS to load data from multiple sources into PDW Server.
  • Managed data workflows for real-time processing, improving data accessibility and ensuring up-to-date information for decision-making.
  • Designed Power BI Reports utilizing cross tabs, scatter plots, pie bar, map charts and density charts, Used DAX queries for computed columns in Power BI.
  • Created Financial Power BI Dashboard for the stake holders and Product owners for business decision making.
  • Wrote complex T-SQL scripts within SSIS Execute SQL tasks to handle multi-step data transformations, aggregations, and partitioning for improved query performance on target systems.
  • Automated data quality checks within SSIS using T-SQL to identify inconsistencies, missing data, or integrity violations before committing changes to target tables.
  • Created dynamic T-SQL code in SSIS to perform conditional logic for data loading and transformation based on external parameters, enabling flexible, reusable ETL workflows.
  • Deployed consistent security frameworks utilizing Azure Key Vault to store secrets and credentials used by Databricks, encrypting data at rest in Azure Data Lake Storage.
  • Led implementation of SDK-driven data validation tools to improve pipeline reliability by 30%.
  • Employed error-handling and data validation techniques across processing workflows, reducing data discrepancies and enhancing data quality.
  • Expert in real-time data processing and analytics using technologies such as Apache Spark, Spark Streaming, Scala, and Kafka, delivering insights that drive significant business improvements.
  • Automated recurring data tasks by creating dynamic stored procedures and complex SQL scripts, reducing manual workload and enabling consistent, timely data processing.
  • Developed many python functions for transforming data to be used within Azure data factory.
  • Designed data extraction routines from multiple sources, utilizing T-SQL for ETL processes, ensuring efficient integration into target databases.
  • Built and maintained release pipelines within Azure DevOps, enabling faster development cycles and minimizing deployment errors.
  • Led the migration of large-scale datasets and database objects from on-premises SQL Server to Azure Synapse, using dedicated SQL pools for optimal data storage and processing performance in the cloud.
  • Migrated complex T-SQL queries, stored procedures, and views from on-prem SQL Server to Azure Synapse, ensuring compatibility and improved performance using Synapse’s distributed architecture.
  • Created automated testing frameworks in DevOps to validate data transformations and pipeline processes, ensuring data accuracy in production environments.
  • Developed and optimized big data pipelines using Cloudera Hadoop Stack, HDFS, and Spark, improving processing efficiency.
  • Implemented environment-specific validations with Git integration, permitting quick rollbacks and thoroughly tested merges before final deployment to the production cluster in Azure Databricks.
  • Designed and Developed Azure Data Factory (ADF) Pipelines to load the data from different sources including Blobs, SharePoint List, Azure Data warehouse, Oracle, DB2, SAP, Hadoop etc.
  • Developed various T-SQL stored procedures, functions, views and triggers.
  • (Client: Schlumberger)

Senior Data Analyst/ Sr. ETL Database Developer

Vonnasys
Houston
09.2018 - 03.2019
  • Developed ETL processes for Oracle to SQL Server and Microsoft AX integration.
  • Loaded and transformed Oracle data into SQL Server for ERP integration.
  • Fetched the Source data from Oracle, perform data cleansing, transformation before loading into the final destination tables.
  • Leveraged T-SQL for data transformations and processing of complex business rules, enabling meaningful insights and streamlined data flows.
  • Created SSIS Packages for each entity that includes Customer Bank Account, Vendor Bank Account, Subscription billing, Sales Partner.
  • Automated workflows and optimized data transformation performance.
  • (Client: Direct Energy)

Senior Data Analyst/ Senior Technical Consultant

Larson and Toubro Infotech
Morristown
09.2018 - 03.2019
  • Developed SQL queries and stored procedures for hierarchical data extraction.
  • Managed claims data extraction, transformation, and loading from Oracle and AS400 systems.
  • Created SSIS packages for claims processing and exception handling.
  • Ensured data accuracy and efficiency in real-time data workflows.
  • (Client: Crum & Forster)

Senior Data Analyst/ Senior Associate-Projects

Cognizant Technology
College Station
02.2017 - 06.2018
  • Led data migration and validation projects for healthcare client databases.
  • Designed and implemented data migration from Oracle to new systems.
  • Optimized large datasets processing, ensuring compliance and security.
  • Developed validation and archival packages for high-volume data.
  • (Client: Nationstar Mortgage, CVS Pharmacy)

Senior Data Analyst

Hexaware Technology
06.2014 - 01.2015
  • Supported production environments with SQL, SSIS, and Unix scripting.

Senior Data Analyst/ Database ETL Engineer

VSquare Infotech
12.2013 - 03.2014
  • Developed ETL processes for data migration and cleansing.
  • Loaded Oracle data into SQL Server for business applications.
  • Created SSIS packages for data transformation and validation.
  • Automated workflows for high-volume data processing.
  • (Client: EY)

Education

Master of Science - Data Science

Regis University
Denver, CO
01.2021

Skills

  • SQL Server and PLSQL
  • PostgreSQL
  • Data integration and transformation
  • Azure Synapse and Data Factory
  • Azure DevOps
  • Databricks and Delta Lake
  • AWS Glue and Lambda
  • PySpark and Spark SQL

Certification

Microsoft Certified Data Engineer Associate, Microsoft, 01/01/25

Immigration Status

US Citizen

Timeline

Lead Data Analyst and Engineer

Infinite/Scan Health
10.2025 - Current

Senior Data Analyst & Data Engineer

American Family Insurance through Cloudpaths/EY
01.2025 - 09.2025

Senior Data Analyst / Data Engineer

Amsive
05.2022 - 11.2024

Data Analyst / Database Developer

Texas Education Agency
03.2020 - 05.2022

Technology Lead – US/Data Analyst

Infosys
03.2019 - 03.2020

Senior Data Analyst/ Sr. ETL Database Developer

Vonnasys
09.2018 - 03.2019

Senior Data Analyst/ Senior Technical Consultant

Larson and Toubro Infotech
09.2018 - 03.2019

Senior Data Analyst/ Senior Associate-Projects

Cognizant Technology
02.2017 - 06.2018

Senior Data Analyst

Hexaware Technology
06.2014 - 01.2015

Senior Data Analyst/ Database ETL Engineer

VSquare Infotech
12.2013 - 03.2014

Master of Science - Data Science

Regis University
Monish Joshy