Data Engineer transitioning from 6+ years of technical product leadership into hands-on engineering. Over the past 2 years, I've built end-to-end ELT pipelines using BigQuery, Databricks, PySpark, Kafka, Snowflake, and dbt — through certifications and portfolio projects spanning cloud warehousing, streaming ingestion, and dimensional modeling.
Before this pivot, I spent 6+ years in enterprise operations and cross-functional delivery at Tata Motors, defining data requirements and translating business needs into technical specifications across cloud data migrations and connected vehicle systems. That experience now shapes how I approach engineering work — stakeholder-aware, delivery-minded, and grounded in real production context.
Currently based in India and open to Data Engineering and Analytics Engineering roles where I can bring both technical depth and stakeholder fluency from my product background. Happy to connect!
Enterprise E-Commerce Data Platform Pipeline (Medallion Architecture)
An end-to-end data engineering lakehouse platform built on top of multi-format enterprise marketplace data shares using Apache Spark, Delta Lake, and Databricks, implementing a comprehensive Three-Tier Medallion Architecture
Tech Stack & Tools: Lakehouse | Databricks | Pipelines | Spark | Visualization
NYC Yellow Taxi Analytics Pipeline
Documents an end-to-end data pipeline that handles real-time stream ingestion and analytics for NYC Yellow Taxi trip data. The primary objective is to transform raw parquet records into structured operational insights.
Tech Stack & Tools: GCP | Terraform | Apache Kafka | Docker | BigQuery | dbt Core | Streamlit | Data Visualization
Superstore Data Analysis Performance Dashboard
Documents the Exploratory Data Analysis (EDA) and analytical layout created for the Superstore Sales dataset. The primary objective is to transform raw sales records into structured operational insights.
Tech Stack & Tools: SQL | Power BI | DAX | Data Visualization