// about // skills // experience // projects // contact
AVAILABLE FOR OPPORTUNITIES

Abhisekh Nayak

Data Engineer & Business Intelligence Analyst

Designing scalable data pipelines and lakehouse architectures that turn raw data into business insight. 4+ years building production-grade solutions on Azure, processing 10M+ records daily.

4+
YEARS EXPERIENCE
10M+
RECORDS / DAY
99.8%
DATA ACCURACY
15+
ETL PIPELINES
SCROLL

Building the data
infrastructure
of tomorrow.

I'm a Data Engineer and BI Analyst based in Bhubaneswar, India, with 4+ years of experience translating complex data challenges into elegant, scalable solutions.

Currently at Tata Consultancy Services, I architect end-to-end data pipelines for TD Bank Canada on Azure — building systems that handle 10M+ records daily with zero downtime and 99.8% accuracy.

My engineering philosophy centers on building reusable, maintainable, production-grade data systems that empower teams through self-service analytics and automated quality frameworks.

⚡
80% Effort Reduction Built a self-service data extraction utility in Databricks that eliminated manual retrieval workflows entirely.
🔬
60% Validation Speedup Engineered automated data quality frameworks now running across 100+ production Delta tables.
📊
35% Query Optimization Reduced MySQL execution time through advanced indexing, normalization, and query restructuring.

Certifications

  • IBM Data Engineering Professional Certificate
  • Snowflake Data Engineering Professional Certificate
  • Meta Database Engineer Professional Certificate
  • Tableau BI Analyst Professional Certificate
  • tcsAI Idea Igniter & tcsAI Spark

Technical
Expertise

Data Engineering
ETL/ELT Pipelines Data Modeling Data Warehousing Data Integration Pipeline Development Delta Lake
Big Data & Processing
Apache Spark PySpark Distributed Computing Databricks
Cloud & Architecture
Microsoft Azure Azure Data Factory ADLS Gen2 Data Lakehouse Databricks
Programming
Python SQL Pandas NumPy SQLAlchemy
Databases
MySQL PostgreSQL SQL Server Oracle DB BigQuery
Analytics & BI
Tableau Power BI Advanced Excel KPI Design Data Visualization
Orchestration & ETL Tools
Apache Airflow Azure Data Factory SSIS
Tools & Platforms
Git / GitHub Jupyter Notebook Confluence DBFS BeautifulSoup Selenium

Work
History

OCT 2024 — PRESENT
Tata Consultancy Services
System Engineer — Cloud Data Engineering & ETL
CLIENT: TD BANK CANADA · BHUBANESWAR, ODISHA
  • Designed and implemented scalable ETL/ELT pipelines using Azure Data Factory and Databricks within a data lakehouse architecture, processing 10M+ records daily with zero downtime.
  • Built distributed data processing workflows using PySpark, enabling efficient transformation and validation of large-scale datasets.
  • Engineered automated data quality frameworks, reducing validation effort by 60% and improving data accuracy to 99.8%.
  • Developed a reusable Delta Table Metadata Extraction framework automating schema, storage, and transaction history analysis across 100+ production tables.
  • Built a self-service data extraction utility in Databricks, reducing manual data retrieval effort by 80%.
  • Optimized SQL queries, partitioning strategies, and compute usage, improving pipeline performance and reducing execution costs.
Azure PySpark Delta Lake Databricks ADF ADLS Gen2
DEC 2021 — JUL 2024
CSM Technologies Pvt. Ltd.
Software Engineer — Data Analytics
BHUBANESWAR, ODISHA
  • Developed 15+ Python-based ETL pipelines and data integration workflows, improving data processing efficiency by 40% across multiple business units.
  • Designed and deployed interactive Tableau dashboards enabling KPI tracking and data-driven decision-making for stakeholders.
  • Optimized MySQL database schemas and complex queries, reducing execution time by 35% through indexing and normalization.
  • Built automated data ingestion pipelines using BeautifulSoup and Python, processing 50K+ data points monthly.
  • Performed exploratory data analysis using Pandas and SQL, identifying trends and generating actionable insights.
Python Tableau MySQL Pandas BeautifulSoup SQL

Featured
Work

01 / 03

Databricks File Extraction Utility

JAN — FEB 2026

A self-service data extraction utility automating retrieval of metadata and files from Azure Data Lake Gen2 within a lakehouse architecture. Supports structured (CSV, JSON, TXT) and binary files (PDF, COPYBOOK) with format-agnostic processing. Reduced manual data retrieval effort by 80%.

Databricks PySpark ADLS Gen2 DBFS Python
02 / 03

Delta Table Metadata Extraction Utility

MAR 2026

Automated metadata extraction framework capturing schema, storage, and transaction history for Delta Lake tables at scale. Generates multi-sheet Excel reports for data governance and auditing. Standardized metadata workflows across datasets, improving consistency and reducing manual analysis effort.

Databricks PySpark Delta Lake Excel SQL
03 / 03

Cricket World Cup Data Ingestion Pipeline

OCT 2023

End-to-end data ingestion pipeline collecting and processing cricket match data from CSV files and live web sources using BeautifulSoup. Automated extraction and transformation workflows stored in MySQL, enabling efficient querying and dynamic points table computation with Pandas and SQL.

Python Pandas MySQL BeautifulSoup Jupyter

Academic
Background

B.Tech in Computer Science & Engineering
Gandhi Institute for Technological Advancements, Bhubaneswar
AUG 2016 — AUG 2020
Data Structures & Algorithms DBMS Data Analytics Machine Learning Cloud Computing Operating Systems Software Engineering
8.21
CGPA / 10

Recognition &
Awards

🧙
Tech Wizard — Python
Recognized at CSM Technologies for excellence in Python development and data processing solutions.
🤖
tcsAI Idea Igniter
Awarded by TCS for active participation and innovation in AI-driven initiatives and projects.
✨
tcsAI Spark
Certified for demonstrated contribution to AI-based problem-solving and data-driven solution development.
🏆
Maitree Prize (2×)
Received twice for outstanding contributions and performance in organizational initiatives at TCS.
💡
TCS AI Hackathon
Participated in internal TCS AI Hackathon contributing to AI-based problem-solving and data-driven solutions.

Let's build
something
great.

Open to data engineering roles, consulting opportunities, and interesting collaborations. Whether you have a project in mind or just want to connect — reach out.