Data Engineer | Azure Databricks | PySpark | Data Pipeline Testing & QA Automation
Send a job offer directly to this candidate
Data Engineer with 3+ years of experience designing, optimizing, and testing ETL/ELT pipelines on the Azure Cloud ecosystem using Azure Databricks, Azure Data Factory, ADLS Gen2, PySpark, and Delta Lake. Proven track record processing 50M+ records/day with 99%+ pipeline uptime, a 35% performance improvement, and 30% reduction in manual deployment effort. Skilled in data validation, schema enforcement, reconciliation testing, Medallion Architecture, Unity Catalog governance, Workflow Orchestration, and CI/CD automation across enterprise data platforms.
Data Engineer - Accenture - Bengaluru, India
(2023-09)
Design, optimize, and test ETL/ELT pipelines on Azure Databricks and Azure Data Factory for enterprise data solutions. Develop PySpark transformations, implement data governance and validation frameworks, and automate CI/CD deployment pipelines. Drive performance optimization and maintain 99%+ pipeline reliability.
Data Engineer - Retail Sales ETL Pipeline - Accenture - Bengaluru, India
(2023-09)
Architected scalable 50M+ daily record pipeline using Databricks, Azure Data Factory (Triggers, Data Flows, Copy Activity, Control Flow), and Delta Lake. Achieved 60% reduction in analyst wait time and 40% processing improvement.
Data Engineer - Customer Data Platform - Accenture - Bengaluru, India
(2023-09)
Built enterprise-grade data lakehouse standardising 10M+ customer records with zero-defect delivery. Implemented SCD Type 1 & 2, dimensional modeling, and Data Governance frameworks for compliance.
Bachelor of Technology (B.Tech) - Electronics & Communication Engineering - Bharath University
Intermediate (12th Grade) - MPC