Vignesan Saravanan

Vignesan Saravanan

Azure Data Engineer

Passionate about building scalable data solutions on Microsoft Azure. Experienced in designing and implementing robust data pipelines, analytics platforms, and cloud-native architectures.

About Me

With 3.5+ years of total work experience as an Azure Data Engineer, I have keen experience in helping companies collect, collate, and explore digital assets. I possess skilled technical expertise in Azure services ranging from Azure Databricks, Azure Data Factory, Microsoft Fabric and cloud services.

I'm a dedicated big data industry professional with a history of meeting company goals utilizing consistent and organized practices. Skilled in working under pressure and adapting to new situations and challenges to enhance the organizational brand.

Professional at building the infrastructure required for optimal ETL/ELT of data from a wide variety of data sources using SQL and Azure. I have the ability to transform complex business requirements into data engineering specifications and extensive experience in developing Data Flows and Pipelines in Microsoft Azure Data Factory and data lakes.

My background includes Data Warehousing and Analytics, with proficiency in AGILE SCRUM methodology and tools such as JIRA. I'm experienced in handling reporting and dashboard services using Big Data Analysis, Power BI, and Power BI Service.

ETL Pipelines Data Warehouse Microsoft Fabric Azure Databricks Data Modelling Agile/Scrum

Quick Facts

Experience: 3.5+ Years
Location: India
Specialization: Azure Data Platform
Languages: Python, SQL, C#

Technical Skills

Azure Services

Azure Data Factory 92%
Azure Databricks 88%
Microsoft Fabric 85%
Azure Data Lake Gen2 90%
Azure DevOps 80%

Programming & Data

SQL 95%
Python 88%
PySpark 85%
Apache Spark 82%
Data Modelling 87%

Tools & Analytics

Power BI 90%
Grafana 75%
ICEDQ A-B Testing 78%
JIRA 85%
Agile/Scrum 88%

Professional Experience

Azure Data Engineer

Harmony Bioscience | USA

Project: Healthcare Patient Admission & Readmission Analytics

Apr 2025 - Present
  • • Designed and developed end-to-end data pipelines in Microsoft Fabric, including data ingestion, transformation, and loading using Lakehouse and delta format
  • • Automated daily ingestion of raw CSV and Excel files from Azure Data Lake Gen2 using Fabric Pipelines to ensure timely data availability
  • • Built transformation logic in Fabric Notebooks using PySpark and SQL to cleanse, enrich, and standardize patient admission and readmission data
  • • Implemented a multi-layered data architecture (bronze → silver → gold) to improve data traceability and optimize reporting performance
  • • Enabled real-time monitoring and reporting by integrating the gold layer data into Power BI, delivering insights into patient flow and departmental efficiency
  • • Ensured data consistency, accuracy, and reliability through schema validations and transformation rules embedded in the notebook flows
  • • Developed mobile-friendly, interactive dashboards to support clinic stakeholders in making data-driven decisions
  • • Scaled the solution for long-term data growth and reduced manual workload for healthcare staff by automating report generation

Azure Data Engineer

PepsiCo | USA

Project: FP&A Data Foundation Team

Nov 2023 - Mar 2025
  • • Developed Databricks notebooks to streamline source-to-silver and silver-to-gold workflows, leveraging PySpark and SQL for efficient data processing
  • • Built Azure Data Factory pipelines encompassing a comprehensive set of activities to facilitate seamless data movement and transformation
  • • Implemented both automated and manual ETL processes efficiently handling large datasets with millions of rows to ensure data accuracy and integrity
  • • Conducted thorough validation of silver and gold layer data using the ICEDQ tool, ensuring the quality and reliability of the processed data
  • • Orchestrated the scheduling and monitoring of Autoloader jobs daily through the ingestion portal, optimizing data loading processes for timely and accurate results
Technologies: PySpark, SQL, Azure Databricks, Azure Data Factory, ADLS Gen2

Azure Data Engineer

Harmony Bioscience | Chennai, India

Project: Sales Dashboard Enhancement

Aug 2022 - Oct 2023
  • • Created multiple reports related to sales, inventory, shipment delivery, and dashboard for Harmony Biosciences business users
  • • Created Azure data factory pipelines with a set of activities for comprehensive data processing workflows
  • • Applied automated and manual ETL processes across millions of rows of data to ensure optimal performance and accuracy
  • • Developed interactive dashboards and reports to support business decision-making processes
Technologies: PySpark, SQL, Power BI, Azure Databricks, Azure Data Factory, ADLS Gen2

Azure Data Engineer

Cricket Social | Chennai, India

Project: CS Data Ingestion & Transformation

May 2022 - Jul 2022
  • • Created Azure data factory pipelines with a comprehensive set of activities for seamless data processing
  • • Applied automated and manual ETL processes across millions of rows of data to ensure data accuracy and performance
  • • Created Synapse SQL Pool tables to store processed results in tabular format for Data Scientists and Data Analysts access
  • • Used PySpark to distribute data processing on large batch datasets, significantly improving ingestion speed and efficiency
  • • Supported implementation and active monitoring of controls and programs for precision and efficacy
Technologies: PySpark, SQL, MS Azure, Databricks, Azure SQL Database, Synapse SQL Pool, ADB Blob/ADLS Gen2, Agile, Scrum

Junior Data Engineer

HexaCorp | Chennai, India

Project: Data Migration

Mar 2022 - May 2022
  • • Assessed on-premises SQL Server databases and identified comprehensive migration requirements
  • • Designed and developed Azure Data Factory pipelines to migrate data from on-premises SQL Server databases to Azure SQL Database
  • • Tested migration pipelines thoroughly to ensure data was migrated accurately and completely without data loss
  • • Deployed migration pipelines and monitored the entire migration process for optimal performance
Technologies: PySpark, SQL, Databricks, Azure Data Factory, Azure SQL Database, Logic App, MS SQL Server

Junior Data Engineer

Fresh Service | Chennai, India

Project: Fresh Service Integration

Jan 2022 - Mar 2022
  • • Created Azure SQL database to load REST API data using Data Factory web activity and copy activity
  • • Created view tables and provided access permissions for Power BI developers to create comprehensive dashboard reports
  • • Developed ticket status reports and inventory asset reports for business stakeholders
  • • Automated reporting processes using Power BI service scheduling concepts for timely data delivery
  • • Implemented data security measures and restricted export options for sensitive reports
Technologies: Azure Data Factory, ADLS Gen2, Azure SQL Database, Power BI, Logic App

Junior Data Analyst

Pizza Pizza | Chennai, India

Project: Pizza Pizza E-commerce Analytics

Dec 2021 - Jan 2022
  • • Developed comprehensive data modeling and sales reporting solutions for pizza ordering e-commerce platform using cloud and BI tools
  • • Followed SDLC methodologies throughout all deployment activities including design, development, testing, and quality assurance support
  • • Created user and admin dashboards with integrated database activities for enhanced business intelligence
  • • Implemented order status monitoring page to provide real-time tracking capabilities for customers and administrators
  • • Designed and developed interactive reports and visualizations to support business decision-making processes
Technologies: MySQL, Power BI Desktop, Power BI Service, GitHub

Featured Projects

Real-time Analytics Platform

Built a comprehensive real-time analytics platform using Azure Event Hubs, Stream Analytics, and Power BI for processing IoT sensor data.

Azure Event Hubs Stream Analytics Power BI

Data Lake Architecture

Designed and implemented a scalable data lake solution on Azure Data Lake Storage Gen2 with automated data governance and cataloging.

ADLS Gen2 Azure Purview Databricks

ML Pipeline Automation

Created an end-to-end ML pipeline using Azure ML and Data Factory for automated model training, validation, and deployment.

Azure ML Data Factory Python

Education

Bachelor of Engineering (Computer Science & Engineering)

Anna University

Comprehensive computer science education with focus on software engineering, data structures, and system design

79.25% Aggregate

Graduated

Key Coursework:

Data Structures & Algorithms Database Management Systems Software Engineering Computer Networks Operating Systems Web Technologies

Higher Secondary Certificate (12th Grade)

State Board of Higher Secondary Education

Science stream with Mathematics, Physics, Chemistry, and Computer Science

85% Aggregate

Secondary School Certificate (10th Grade)

State Board of Secondary Education

Strong foundation in core subjects with excellent academic performance

87.8% Aggregate

Academic Excellence

79.25%

Engineering Degree

Computer Science & Engineering

85%

Higher Secondary

Science Stream

87.8%

Secondary School

All Subjects

Consistent academic excellence throughout educational journey with strong foundation in computer science fundamentals

Certifications

Microsoft Azure Data Engineer Associate Certification

Certification Code: DP-203

Advanced data engineering skills on Microsoft Azure platform

Microsoft Azure Data Fundamentals Certification

Certification Code: DP-900

Core data concepts and Azure data services fundamentals

Microsoft Azure Fundamentals Certification

Certification Code: AZ-900

Foundational knowledge of cloud services and Microsoft Azure

Get In Touch

Let's Connect

I'm always interested in discussing new opportunities, data engineering challenges, or potential collaborations. Feel free to reach out!

vigneshsiva3699@gmail.com

Email

+91 7448974705

Phone

LinkedIn Profile

Professional Network

YouTube Channel

Tech Content