OPEN TO DATA ENGINEERING OPPORTUNITIES

Engineering the
flow of intelligence.

Data Engineer with 4+ years building cloud-scale data platforms, resilient ETL/ELT pipelines, and analytics-ready lakehouses.

LIVE PLATFORM BLUEPRINT RUNNING
01IngestAPIs · CDC · Files
02StreamKafka · Kinesis
03TransformSpark · dbt
04LoadLakehouse · DWH
ORCHESTRATED BY AIRFLOWQUALITY GATES ACTIVELINEAGE CAPTURED
4+Years building
data products
55+Enterprise systems
integrated
3Major cloud
platforms
A 3D visualization of a data engineer working with cloud data pipelines
ORCHESTRATIONCloud-native
Pipeline healthy24 jobs running
SCROLL TO EXPLORE

01 / EXPERIENCE

Turning raw data into
reliable decisions.

From real-time financial reporting to customer 360 platforms, I design production pipelines that teams can trust.

MAY 2025 — JUN 2026
FR

Data Engineer

Federal Reserve Bank of Cleveland Ohio, USA

  • Integrated 20+ financial, economic, and operational source systems into AWS S3 and Redshift using Glue, PySpark, SQL, and Airflow.
  • Built batch and event-driven ingestion with Lambda, Kinesis, Kafka, and REST APIs, reducing data availability to hourly refreshes.
  • Developed reusable dbt and Great Expectations frameworks for schema validation, monitoring, and data quality.
AWS GluePySparkAirflowRedshiftdbt
SEP 2022 — AUG 2024
AB

Azure Data Engineer

Aditya Birla Capital Mumbai, India

  • Built a Customer 360 platform from 25+ systems using ADF, Databricks, ADLS Gen2, and SSIS.
  • Designed Bronze, Silver, and Gold Delta Lake layers for 10M+ records using CDC and incremental loading.
  • Implemented governance with Purview, Unity Catalog, Key Vault, RBAC, and lineage across shared platforms.
Azure Data FactoryDatabricksDelta LakeSynapseSnowflake
JUL 2021 — AUG 2022
TC

Data Engineer

Tata Communications Hyderabad, India

  • Built Python, SSIS, Hadoop, and Hive workflows across telecom and healthcare datasets.
  • Created reusable datasets from SQL Server, PostgreSQL, MySQL, and MongoDB with trusted validation layers.
  • Supported 100+ business users with analytics-ready models for Power BI, Tableau, Looker, and Excel.
PythonSQL ServerHadoopHivePower BI

02 / SELECTED PROJECT

BigQuery, meet
business questions.

Request a walkthrough
FEATURED BUILD

AI BigQuery
Analytics Assistant

A cloud-native retail analytics application that combines automated ETL, warehouse-grade modeling, interactive dashboards, and natural-language insight discovery.

GCPCloud platform
RAGAI-enhanced discovery
CI/CDAutomated deployment
BigQueryStreamlitCloud RunPythonPlotly
LIVE DATA FLOW
CSV
Sources
ETL
Validate + Model
BQ
Warehouse

03 / TECHNICAL TOOLKIT

Tools for every
stage of the flow.

01

Cloud platforms

AWS, Azure, and GCP architectures designed for scale, security, and operational clarity.

aws▣ AzureGCP
02

Processing & storage

Lakehouse and warehouse foundations built for high-volume batch and real-time workloads.

Apache SparkDatabricksSnowflakeKafka
03

Engineering & trust

Orchestration, testing, governance, and delivery practices that make pipelines dependable.

AirflowdbtTerraformDocker
04

AI & analytics

Modern data products, from semantic models to RAG-enabled data discovery.

PythonRAGPower BIMLflow

04 / CREDENTIAL

DATABRICKS

Certified Data Engineer
Professional

Credential ID: 191605188

Verify credential

05 / EDUCATION

Master of Engineering
in Computer Science

University of Cincinnati Ohio, USA

AUG 2024 — APR 2026

06 / CONTACT

Let’s build the next
data advantage.

Have a data platform challenge, a role, or an idea? I’d love to hear about it.

gongatisandeep.dev@gmail.com
+1 513 551 0961UNITED STATES