Data Engineer | PySpark | AWS Glue Apache Airflow | SQL
Send a job offer directly to this candidate
Data Engineer with 2.6 years of experience in designing, developing, and migrating scalable ETL pipelines using PySpark and AWS Glue in the insurance domain. Experienced in migrating legacy BDM workflows to AWS Glue and converting TWS schedulers into Apache Airflow DAGs for workflow orchestration and monitoring. Skilled in building configuration-driven ETL frameworks, batch-based incremental processing, data validation, and source-to-target reconciliation for large-scale claims data processing.
Hands-on experience with PySpark transformations, Apache Hive, Amazon S3, PostgreSQL, and AWS services to support scalable cloud-based data engineering solutions. Proficient in CI/CD deployments using Azure DevOps and GitHub Actions, with expertise in workflow monitoring through Airflow, Control-M, and CloudWatch.
Data Engineer at Tata Consultancy Services (TCS) (2024-02 – Present)
Project: Upstream Rebuild – Data Migration & ETL Modernization (Aviva UK Insurance Domain). Source Systems: Ireland Claims, Pelican Claims, Guidewire Claims. Tech Stack: PySpark, AWS Glue, Apache Airflow, MWAA, SQL, Apache Hive, Amazon S3, PostgreSQL, CloudWatch, Azure DevOps, GitHub Actions, Control-M
Bachelor of Engineering (B.E.) in Electronics & Telecommunication Engineering – RMK Engineering College (2019 – 2023)