Data Engineer | PySpark | SQL | Python | AWS | Azure | Snowflake | Databricks | Airflow
Send a job offer directly to this candidate
Data Engineer with 4+ years of experience designing, developing, and optimizing scalable ETL/ELT pipelines and cloud-native data platforms. Skilled in Python, SQL, PySpark, Apache Spark, Databricks, Azure Data Factory, Snowflake, and DBT, with hands-on expertise in data ingestion, transformation, Delta Lake, Medallion Architecture, and workflow orchestration across Azure and AWS. Experienced in processing structured and semi-structured data, implementing data quality, governance, and security best practices, and delivering analytics-ready datasets for BI and advanced analytics.
Proven ability to build secure, high-performance data engineering solutions using Azure Synapse Analytics, ADLS Gen2, AWS S3, AWS Glue, Redshift, Apache Airflow, Docker, Git, CI/CD, and Agile methodologies.
Data Engineer at CGI (2023-08 – Present)
Healthcare Data Integration Platform for a Leading Global Pharmaceutical & Clinical Research Organization. Designed and developed an enterprise-scale clinical data integration platform on AWS to ingest, transform, and manage clinical trial data from CTMS, EDC, SQL Server, and laboratory systems. Built scalable ETL pipelines using AWS Glue and PySpark, implemented Medallion Architecture, orchestrated workflows with Apache Airflow, and enabled Snowflake analytics through Snowpipe.
Data Engineer at CGI (2023-08 – Present)
Customer 360 Analytics Platform for a Global Research & Advisory Organization. Developed an enterprise Customer 360 Analytics Platform on Microsoft Azure using ADF, Databricks, Delta Lake, and ADLS Gen2 to deliver curated datasets for retention analytics and Power BI reporting.
Data Engineering Intern at Icode Technology (2022-08 – 2023-08)
B.Tech – G.H. Raisoni College of Engineering (2019 – 2023)