Data Engineer | Bengaluru | Open to opportunities

Building pipelines that move data. with governance.

5+ years building scalable data pipelines and governance systems across Databricks, DBT, GCP and Collibra. From ELT at scale to metadata catalogues that business users actually use.

5+
Years experience
2x
Databricks certified
3
Companies

About

Data systems built for longevity

The pipeline which runs when none is awake.

I'm a Data Engineer with a track record across financial services, FMCG, and banking. My work sits at the intersection of raw engineering - ELT jobs, transformation of raw data via DBT and its DAGs, Databricks SQL editors to test it and the less-glamorous-but-essential layer of data governance: making sure the right people see the right data and understand where it came from.

At Tyson Foods I built a Python application that compared data assets between GCP Data Catalog and Collibra, automatically raised JIRA tickets for discrepancies, and designed custom governance workflows in Collibra's BPMN engine. Also added GCP policy tags and IAM roles for data access control.
Currently at Iksula, I'm leading a legacy SQL Server → Databricks migration for Nationwide Building Society.

Engineering depth

Multi-layer dbt models (staging -> intermediate -> mart), incremental load strategies, and DAG optimization to reduce runtime.

Governance mindset

Collibra workflow design, metadata lineage, REST API integrations between cloud and governance tool.

Experience

Where I've built things

From retirement domain to food manufacturing and back to setting up Cloud infrastructure.

Iksula Private Limited Apr 2025 - Present
Data Engineer
  • Designed dbt models across staging, intermediate, and mart layers for Nationwide Building Society, standardizing raw enterprise data into analytics-ready datasets.
  • Led migration of legacy Microsoft SQL Server infrastructure to Databricks, improving processing efficiency and enabling scalable transformations.
  • Leveraged Jinja macros and dbt tests to eliminate SQL duplication, enforce schema contracts, and reduce DAG execution time.
  • Built PySpark transformation pipelines in Databricks for large-scale datasets with schema validation and DQ checks baked in.
DatabricksDBTPySpark SQL ServerData MigrationGIT Jobs and Pipelines
Tyson Foods May 2023 - Mar 2025
Associate Developer - Cloud Platforms
  • Designed and optimized scalable dbt ELT models with incremental strategies according to industry standards to serve business needs and analytics.
  • Built and owned a Python application comparing data assets between GCP Data Catalog and Collibra - auto-detecting discrepancies and raising JIRA tickets without manual intervention.
  • Solely managed development of custom Collibra workflows (BPMN) enabling business units to classify and govern data assets end-to-end.
  • Created GCP policy tags, IAM roles, and security configurations for a high-priority data access control project.
GCPCollibradbtGIT PythonBPMNJIRA APIData Catalog
Fidelity Information Services Jun 2021 - May 2023
Associate Developer
  • Implemented the Smart Retirement tool - a mathematical tool predicting optimal retirement cities for participants based on their corpus, using Python.
  • Built the Digital Payments feature enabling direct retirement corpus disbursement to bank accounts.
PythonMathematical ModelingFintechGIT

Skills

Technical toolkit

Tools I reach for in production, and learning to upskill in.

Data Engineering

Databricks DBT PySpark Apache Kafka(Learning) Airflow(Learning) ELT / ETL

Cloud Platforms

GCP BigQuery Azure Databricks Terraform

Data Governance

Collibra GCP Data Catalog BPMN 2.0 Metadata Lineage REST APIs

Languages

Python SQL Groovy Go (learning)

AI & GenAI

Claude API GenAI fundamentals LLM integration

Tools & Practices

GIT JIRA MySQL FastAPI Docker

Certifications

🏅
Databricks Certified Data Engineer Associate
Databricks
🏆
Databricks Certified Data Engineer Professional
Databricks

Projects

Side projects & internal tools

Built in spare time, for real problems.

Side Project - In Production

FairShare

Full-stack group expense splitting app. Handles shared costs, per-person balances, and settlement suggestions. Built with Vite + React, FastAPI, and Supabase.

ReactFastAPISupabasePython
Internal Tool

GCP ↔ Collibra Sync

Python application that compares data asset inventories between GCP Data Catalog and Collibra, detects mismatches, and automatically raises JIRA tickets for discrepancies.

PythonGCPCollibra APIJIRA API
Side Project - POC

GrowwLens

CAMS mutual fund PDF parser and portfolio analytics tool. Parses transaction statements using pdfplumber, enriches with AMFI NAV data, and surfaces portfolio analytics via a FastAPI + React dashboard.

FastAPIpdfplumberAMFI APIReact

Contact

Let's build something solid

Open to Data Engineering roles. Currently based in Bengaluru, interested in relocation.

Whether it's a pipeline problem, a governance challenge, or a new role - reach out.