Site Reliability Engineer | Devops Engineer
Send a job offer directly to this candidate
9+ years of experience in IT roles including DevOps, Site Reliability Engineer, Cloud Engineer, and Build and Release Engineer, specializing in automation, integration, deployment, and maintenance on Unix/Linux/VM platforms. Extensive expertise in AWS Cloud Services (EC2, VPC, EBS, RDS, CloudWatch, CloudFormation, IAM, S3), managing full lifecycle, automation, and security for scalable, high-availability environments. Proficient in CI/CD tools for automating build, test, and deployment pipelines, significantly improving development workflows and delivery speed.
Skilled in Terraform for Infrastructure as Code (IaC), automating provisioning and management of AWS infrastructure, and ensuring consistency across environments. Strong experience with Ansible for configuration management, including developing playbooks to streamline system configuration and compliance across environments. Expertise in Docker and containerization technologies, including managing Docker images, registries, Compose, and Swarm for scalable microservices.
Managed and optimized Kubernetes clusters for container orchestration, including implementing auto-scaling, network policies, and rolling updates for highly available deployments. Hands-on experience with OpenShift, including configuring CI/CD pipelines, scaling applications, and managing security and deployment across environments. Security-focused DevOps practices, including securing AWS resources, implementing role-based access control (RBAC) in Kubernetes, and using security scanning tools in CI/CD pipelines.
Objectives (SLOs), Service Level Indicators (SLIs), and SLAs to ensure system reliability and proactively monitor application performance. Automated and optimized incident response workflows, significantly improving Mean Time to Recovery (MTTR) and reducing operational overhead. Designed and maintained observability solutions using AWS CloudWatch, Prometheus, and Grafana, enabling proactive monitoring, alerting, and performance optimization.
Led post-incident reviews to identify root causes, implement corrective actions, and drive continuous improvements in system reliability and operational efficiency. Strong experience in version control with Git, Bitbucket, and SVN, ensuring smooth branching, merging, and collaboration across development teams. Focused on scalability, security, and performance optimization across cloud platforms, containers, and Kubernetes/OpenShift environments, ensuring robust infrastructure for critical applications.
Site Reliability Engineer - JPMorgan Chase - Plano, TX
(2024-12)
Site Reliability Engineer - Emirates Airline - Dallas, TX
(2021-09 - 2024-11)
DevOps Engineer - SAP Concur - Bellevue, WA
(2021-01 - 2021-04)
Sr. DevOps Engineer - Capgemini - Atlanta, GA
(2018-05 - 2020-12)