About
Results-driven Data Professional with 2+ years of experience in data analytics and engineering, with strong proficiency in Python, SQL, and Power BI. Experienced in developing end-to-end ETL pipelines for data extraction, transformation, automation, and reporting, along with creating 100+ Power BI dashboards for business insights. Hands-on experience with technologies including PySpark, Kafka, Debezium, AWS Glue, Apache Airflow, PostgreSQL, Snowflake, Docker, and Concourse through professional and internship experience. Applied LLM and Retrieval-Augmented Generation (RAG) techniques in projects to enhance data extraction and processing workflows. Strong ability to understand business requirements and translate data into reliable, actionable solutions.
Skills & Expertise (30)
Work Experience
IT Executive
Kalpataru Projects International Limited
Nov 2024 - Present
Designed, built, and automated end-to-end ETL processes using Python (web scraping), SQL, RPA, and LLM-based workflows to extract, clean, transform, and load data from multiple structured and unstructured sources into reporting-ready formats. Developed 100+ interactive Power BI dashboards used across business teams, translating complex, multi-source datasets into clear, actionable insights that support day-to-day operational and strategic decisions. Built and optimized SQL queries, views, and stored procedures to support fast, reliable analytics, ad-hoc reporting, and high-volume data retrieval. Integrated LLM-based workflows into existing RPA pipelines to automatically parse and structure unstructured documents (PDFs, web content), reducing manual data entry and extraction effort. Partnered directly with business stakeholders to gather requirements, translate them into technical and pipeline specifications, and led User Acceptance Testing (UAT) to validate deliverables before go-live.
Data Engineer Intern
Go Digital Technology
Jun 2024 - Aug 2024
Designed and deployed end-to-end ETL pipelines in Python to automate data processing workflows using a mix of cloud-native and open-source technologies. Built a real-time streaming architecture using Kafka and Debezium to capture change data (CDC) and feed it into downstream relational stores for near real-time analysis. Automated data ingestion and web scraping workflows using CI tooling, with structured, query-ready storage in PostgreSQL. Performed scalable, distributed data transformations using Spark Streaming on cloud infrastructure, supporting large-volume data processing needs.
Education
MBA (Finance) - NMIMS
- ยท Afghanistan
B.E. (EXTC) - Rajiv Gandhi Institute of Technology
- ยท Afghanistan
Certifications
No certifications added yet
Interested in this developer?
Profile Score Breakdown
Profile Overview
Availability Details
Visa Status
Citizen
Relocation
Depends on Offer
Skills (30)
Click a skill to find developers with the same skill