Data Engineer
Send a job offer directly to this candidate
Data Engineer with hands-on experience designing, building, and optimizing production ETL/ELT pipelines, analytical lakehouses, and cloud data platforms on AWS. Specialized in Python (PySpark, Pandas), Advanced SQL, dbt Core, and Apache Airflow. Demonstrated success processing enterprise batch datasets (~500GB+ scale) and real-time streams (50,000 events/sec), building CDC streaming pipelines (Kafka, Debezium, sub-10ms Redis retrieval), eliminating PySpark join memory skew via key-salting, and deploying automated schema contracts (Pandera) to maintain a 99.9% data freshness SLA.
Data Engineer Consultant at Kaival Technologies (2023-10 – Present)
MSc in Data Science & Analytics with Advanced Research – University of Hertfordshire (2021 – 2023)
BTech in Computer Science – Marwadi University (2016 – 2020)