Back to Developers
nidhi dongre

nidhi dongre

Data Engineer

Pune, Maharashtra 1+ yrs exp 83 ยท Excellent

About

Data Engineer with 1+ year of experience designing and developing scalable ETL/ELT pipelines, data ingestion workflows, data lake solutions, and analytics-ready datasets using Python, SQL, PySpark, Apache Spark, Databricks, AWS S3, and EMR. Hands-on with batch processing, Delta Lake, data quality, data modeling, data warehousing, Spark performance tuning, incremental processing, workflow orchestration, and cloud data platforms. Experienced in building reliable pipelines for structured and semi-structured data and delivering reusable datasets for analytics and reporting.

Skills & Expertise (34)

DBT Intermediate
6.5/10
1
Years Exp
GCP BigQuery Intermediate
6.5/10
1
Years Exp
Data Warehousing Intermediate
6.5/10
1
Years Exp
Star Schema Intermediate
6.5/10
1
Years Exp
SCD Type 2 Intermediate
6.5/10
1
Years Exp
Data Quality Intermediate
6.5/10
1
Years Exp
Data Validation Intermediate
6.5/10
1
Years Exp
Apache Airflow Intermediate
6.5/10
1
Years Exp
Apache Kafka Intermediate
6.5/10
1
Years Exp
SQL Intermediate
6.5/10
1
Years Exp
Git Intermediate
6.5/10
1
Years Exp
GitHub Intermediate
6.5/10
1
Years Exp
Docker Intermediate
6.5/10
1
Years Exp
MySql Intermediate
6.5/10
1
Years Exp
Postgresql Intermediate
6.5/10
1
Years Exp
Jupyter Notebook Intermediate
6.5/10
1
Years Exp
VS Code Intermediate
6.5/10
1
Years Exp
Data Transformation Intermediate
6.5/10
1
Years Exp
Python Intermediate
6.5/10
1
Years Exp
PySpark Intermediate
6.5/10
1
Years Exp
Apache Spark Intermediate
6.5/10
1
Years Exp
ETL Intermediate
6.5/10
1
Years Exp
Data Pipelines Intermediate
6.5/10
1
Years Exp
Data Ingestion Intermediate
6.5/10
1
Years Exp
Batch Processing Intermediate
6.5/10
1
Years Exp
Databricks Intermediate
6.5/10
1
Years Exp
AWS S3 Intermediate
6.5/10
1
Years Exp
EMR Intermediate
6.5/10
1
Years Exp
Azure Data Factory Intermediate
6.5/10
1
Years Exp
ADLS Gen2 Incremental Loads ELT Delta Lake Spark SQL

Work Experience

Data Engineer

2GBR Software Private Limited

May 2025 - Present

Built and maintained 10+ ETL/ELT data pipelines using PySpark, Apache Spark, Databricks, and AWS S3 for structured and semi-structured data, improving pipeline throughput by 30%. Optimized Spark and Spark SQL workloads through partitioning, broadcast joins, predicate pushdown, caching, and query tuning, reducing average job runtime by 25% from approximately 90 minutes to 65 minutes. Designed AWS S3-based data lake workflows processing 30โ€“50 GB daily and delivered cleansed, analytics-ready datasets for reporting and downstream data consumption. Implemented schema validation, deduplication, transformation rules, and automated data-quality checks across more than 5 million records daily, maintaining 99%+ data accuracy. Automated pipeline scheduling, dependency management, retries, monitoring, and failure handling using Apache Airflow; developed SQL/Spark SQL transformations and worked with AWS EMR, Delta Lake, Hive, and HDFS for distributed data processing.

Education

B.Tech โ€“ Artificial Intelligence and Data Science - S.B. Jain Institute of Technology, Management & Research

2021 - 2025 ยท Afghanistan

Certifications

No certifications added yet

Interested in this developer?

Profile Score Breakdown

๐Ÿ“ท Photo 10/10
๐Ÿ“„ Resume 10/10
๐Ÿ’ผ Job Title 10/10
โœ๏ธ Bio 10/10
๐Ÿ› ๏ธ Skills 20/20
๐ŸŽ“ Education 10/10
โฑ๏ธ Experience 8/15
๐Ÿ’ฐ Rate 0/5
๐Ÿ† Certs 0/5
โœ… Verified 5/5
Total Score 83/100

Profile Overview

Member sinceOct 2026