Technical Lead – DevOps Engineer at HCL Technologies Ltd. (2024-05 – 2026-03)
Project: Thryve Digital Health LLP (Healthcare Domain)
- Coordinated the end-to-end application delivery flow — Git source changes, Jenkins builds, deployment, and release validation — across 4 environments (Dev/QA/UAT/Prod), cutting release cycle time by 20%.
- Configured and maintained Jenkins jobs and build agents; resolved build failures and deployment errors, improving pipeline success rate to 95%.
- Managed Git branching, merging, and release tagging strategy for 15+ active repositories, ensuring the correct approved source version reached production every release.
- Built and maintained Docker images for 20+ microservices, standardizing packaging and reducing environment-related deployment issues by 30%.
- Managed Kubernetes Deployments, Services, ConfigMaps, and Secrets across multiple clusters running 50+ pods; verified pod health post-change and remediated unhealthy workloads.
- Executed Kubernetes rolling updates and rollbacks for 10+ releases per month with minimal downtime, validating application health after each change.
- Supported AWS environments (EC2, IAM, S3, VPC, Auto Scaling, Load Balancers) for 15+ applications, performing infrastructure checks and post-change validation.
- Provisioned and maintained AWS infrastructure using Terraform/CloudFormation, standardizing environment setup and reducing manual provisioning errors.
- Diagnosed AWS environment issues by reviewing infrastructure status and configuration changes, reducing mean time to resolution (MTTR) to under 2 hours.
- Authored 25+ Ansible playbooks for repeatable server configuration, cutting manual configuration effort by 40%.
- Wrote Python and Shell scripts to automate recurring operational checks and log analysis, saving approximately 8 hours/week of manual effort.
- Performed Linux/Windows administration — patching, health checks, service restarts, log analysis — across 30+ servers with zero missed maintenance windows.
- Used Prometheus and Grafana dashboards to monitor infrastructure/application health, triaging 20+ alerts per month and escalating to L3 teams as needed.
- Supported 10+ production releases per quarter through pre-change checks, implementation coordination, and post-change validation with a 98% success rate.
- Managed 30+ incidents/service requests per month, analyzing logs and environment health, coordinating technical teams, and documenting resolutions within SLA.
Technical Lead – DevOps Engineer at HCL Technologies Ltd. (2024-05 – 2026-03)
Project: Intel (Semiconductor Domain – DevOps & Release Management)
- Coordinated Git, Jenkins, Docker, and release validation activities to support 8+ application releases per quarter.
- Managed Git branch/merge/tag strategy to maintain correct source versions for controlled deployments across 10+ repositories.
- Investigated Kubernetes deployment issues (pod status, config, application behavior), resolving 90% of incidents without escalation.
- Automated recurring operational tasks using Ansible and Shell scripting, reducing manual execution time by 35%.
- Partnered with development, infrastructure, and support teams during releases to resolve build, deployment, and production issues.
Consultant – DevOps Engineer at Virtusa Consulting Services Pvt. Ltd. (2021-02 – 2024-04)
Client: Citi Bank – GPA Compliance Project (Banking Domain)
- Performed daily health checks, validation, and monitoring for Citi Screening production applications supporting a global compliance user base.
- Monitored AWS-based application environments for availability and post-change behavior across 5+ environments.
- Executed controlled production maintenance (OS VTM, DB PSU, Drop Partition) within approved windows with zero unplanned downtime.
- Raised and implemented 100+ UAT/Production Change Requests, coordinating prerequisites and post-change validation.
- Coordinated application version upgrades and planned releases with Development, Infrastructure, and Solution Architecture teams.
- Performed L2 troubleshooting for production issues, isolating root cause and routing to the appropriate technical team, resolving 85% within SLA.
- Led production incidents from investigation through resolution — analysis, coordination, tracking, and status communication — for 50+ critical incidents.
- Validated application availability and server health post-maintenance across 15+ planned changes per month.
- Identified repetitive operational tasks and delivered automation/scripting solutions, reducing manual effort by 25%.
- Participated in daily client huddles on production issues, planned changes, and cross-team coordination.
- Maintained controlled ITIL processes for incident, change, request, release, and major incident management in a regulated banking environment.