Back to Developers
Daksh Goel

Daksh Goel

Data Engineer

Pune, Maharashtra 1+ yrs exp 84 ยท Excellent

About

Data Engineer with hands on experience of building scalable data pipelines and cloud-based data solutions using Azure Databricks, PySpark, and Delta Lake. Experienced in Data Warehousing, Data Modeling, and end-to-end ETL/ELT pipelines, with hands-on knowledge of Medallion Architecture and Lakehouse design. Focused on developing reliable, efficient, and analytics-ready data solutions.

Skills & Expertise (21)

Azure Databricks Advanced
9.5/10
2
Years Exp
PySpark Advanced
9.2/10
2
Years Exp
Apache Spark Advanced
9.0/10
2
Years Exp
Delta Lake Advanced
9.0/10
2
Years Exp
Spark SQL Advanced
9.0/10
2
Years Exp
Azure Data Lake Storage Advanced
8.5/10
2
Years Exp
SQL Advanced
8.5/10
2
Years Exp
Python Advanced
8.5/10
2
Years Exp
Data Quality Validation Advanced
8.5/10
2
Years Exp
Data Warehousing Advanced
8.5/10
2
Years Exp
Azure Data Factory Advanced
8.0/10
2
Years Exp
Source-to-Target Reconciliation Advanced
8.0/10
2
Years Exp
Azure Blob Storage Advanced
8.0/10
2
Years Exp
Data flows Advanced
8.0/10
2
Years Exp
ETL Advanced
8.0/10
2
Years Exp
Parquet Advanced
8.0/10
2
Years Exp
JSON Advanced
8.0/10
2
Years Exp
CSV Advanced
8.0/10
2
Years Exp
Git Advanced
8.0/10
2
Years Exp
GitHub Advanced
8.0/10
2
Years Exp
Dimensional Data Modelling Intermediate
7.5/10
2
Years Exp

Work Experience

Software Engineer Level 2

Hexaware Technologies

Dec 2024 - Present

Managed 15+ production Databricks pipelines processing 50GB+ daily data across Bronze and Silver layers in a Medallion Architecture. Built data quality validation scripts in SQL and PySpark across 1M+ records (null checks, duplicate checks, source-target row-count match) to ensure data integrity. Diagnosed and resolved 50+ Spark job failures via log analysis (out-of-memory, data skew, schema mismatch), maintaining 99% pipeline SLA. Optimized Delta Lake tables using OPTIMIZE (with Z-ORDER) and VACUUM on a 2TB+ warehouse, improving query performance and reducing storage cost. Supported schema modifications on 5+ core tables (column additions, data-type changes) under senior guidance, using Delta schema evolution. Implemented incremental data processing and transformation logic in Databricks using Auto Loader and Spark Streaming Dataframe to efficiently process large datasets and reduce unnecessary data reprocessing.

Education

B.Tech in Computer Science and Engineering - Graphic Era Deemed to be University

2020 - 2024 ยท Afghanistan

CBSE Class XII - Scottish International School

2018 - 2019 ยท Afghanistan

ICSE Class X - St. Francis School

2017 - 2018 ยท Afghanistan

Certifications

No certifications added yet

Interested in this developer?

Profile Score Breakdown

๐Ÿ“ท Photo 10/10
๐Ÿ“„ Resume 10/10
๐Ÿ’ผ Job Title 10/10
โœ๏ธ Bio 10/10
๐Ÿ› ๏ธ Skills 20/20
๐ŸŽ“ Education 10/10
โฑ๏ธ Experience 9/15
๐Ÿ’ฐ Rate 0/5
๐Ÿ† Certs 0/5
โœ… Verified 5/5
Total Score 84/100

Profile Overview

Member sinceSep 2026

Availability Details

Visa Status

Need Sponsorship

Relocation

Open to Relocation