Back to Developers
Shivam Singh

Shivam Singh

Data Engineer

Noida, India 0+ yrs exp 86 · Excellent

About

Data Engineering fresher with 8 months of internship experience at Lumiq, building data pipelines processing 1M+ records daily, cutting data processing time by 35% using Apache Spark, and achieving 99.9% pipeline accuracy. Skilled in Python, SQL, Apache Airflow, Amazon Redshift, DAX, and Power BI dashboarding, data visualization, and stakeholder reporting. Targeting Data Engineer, Data Analyst, and BI Developer roles.

Skills & Expertise (32)

ETL Pipeline Development Intermediate
7.4/10
0.67
Years Exp
Apache Spark Intermediate
7.0/10
0.67
Years Exp
Apache Airflow Intermediate
7.0/10
0.67
Years Exp
Python Intermediate
6.8/10
0.67
Years Exp
NumPy Intermediate
6.8/10
0.67
Years Exp
Pandas Intermediate
6.8/10
0.67
Years Exp
AWS EC2 Intermediate
6.8/10
0.67
Years Exp
AWS S3 Intermediate
6.8/10
0.67
Years Exp
Amazon Redshift Intermediate
6.8/10
0.67
Years Exp
SQL Intermediate
6.6/10
0.67
Years Exp
MySql Intermediate
6.6/10
0.67
Years Exp
Postgresql Intermediate
6.6/10
0.67
Years Exp
Data Warehousing Intermediate
6.2/10
0.67
Years Exp
Microsoft Excel Intermediate
6.0/10
0.67
Years Exp
Git Intermediate
6.0/10
0.67
Years Exp
Jupyter Notebook Intermediate
6.0/10
0.67
Years Exp
Seaborn Intermediate
6.0/10
0.67
Years Exp
scikit-learn Intermediate
6.0/10
0.67
Years Exp
Matplotlib Intermediate
6.0/10
0.67
Years Exp
TensorFlow Beginner
5.4/10
0.67
Years Exp
Power BI Beginner
5.4/10
0.67
Years Exp
DAX Beginner
5.0/10
0.67
Years Exp
Business Intelligence Beginner
5.0/10
0.67
Years Exp
Data Visualization Beginner
5.0/10
0.67
Years Exp
Dashboarding Beginner
5.0/10
0.67
Years Exp
Stakeholder Reporting Beginner
5.0/10
0.67
Years Exp
Java Apache Hadoop Data Modeling Batch Processing Data Cleaning Data Quality & Validation

Work Experience

Data Engineering Intern

Lumiq

Jun 2025 - Jan 2026

Built 12+ automated ETL pipelines using Python and Apache Airflow, cutting manual data ingestion time by 40% and scaling throughput to 1M+ records/day. Processed 50GB+ datasets per pipeline run with Apache Spark, shrinking job execution time from 4 hours to 85 minutes (a 65% cut). Migrated 3 reporting workflows to Amazon Redshift with partition pruning and sort keys, delivering 2x faster query response times for 5+ BI analysts. Rewrote 20+ slow SQL queries across PostgreSQL and MySQL using CTEs and indexes, trimming average query execution time by 30% (from 10s to 7s). Deployed data validation framework covering 15+ pipeline checkpoints, maintaining 99.9% data accuracy across all downstream reporting systems.

Education

B.Tech – Information Technology - Galgotias University

2022 - 2026 · Afghanistan

Certifications

Database Programming with SQL

Oracle Academy · 2024

Programming Fundamentals using Python

Infosys Springboard · 2024

Data Fundamentals

IBM SkillsBuild · 2023

Interested in this developer?

Profile Score Breakdown

📷 Photo 10/10
📄 Resume 10/10
💼 Job Title 10/10
✍️ Bio 10/10
🛠️ Skills 20/20
🎓 Education 10/10
⏱️ Experience 6/15
💰 Rate 0/5
🏆 Certs 5/5
Verified 5/5
Total Score 86/100

Profile Overview

Member sinceAug 2026