DevOps Engineer
Descrição da vaga
About The Opportunity
We are looking for a Senior DevOps to own the stability, reliability, and continuous improvement of our applications, delivery pipelines, and environments.
This role combines platform engineering with hands-on operational ownership — you'll be the engineer who not only keeps things running, but also builds the tooling, standards, and automation that reduce recurring issues in the first place. Production support and incident response are part of the job, but so is architecting the CI/CD platform, governance, and observability that make that support sustainable.
This role is ideal for someone who thinks like a platform engineer but is comfortable getting hands-on with triage, root-cause investigation, and driving problems through to permanent resolution — not just closing tickets, but fixing the systemic issues behind them.
Responsibilities
- Own application, infrastructure, and CI/CD pipeline reliability end-to-end.
- Handle operational incidents, escalations, and service requests, driving root-cause analysis and permanent fixes — not just workarounds.
- Design, standardize, and evolve CI/CD pipelines and reusable templates for fast, predictable, secure releases.
- Manage and standardize configurations, environments, and infrastructure-as-code practices.
- Support, monitor, and optimize cloud environments (AWS/Azure/GCP), including observability stack ownership (metrics, logs, tracing, alerting).
- Administer and govern DevOps tooling platforms (e.g., GitHub/GitLab/Azure DevOps) — org structure, security policies, access controls, and compliance workflows.
- Automate operational tasks and build internal tools or self-service platform capabilities to reduce manual toil.
- Collaborate with engineering teams to improve system reliability, reduce recurring operational issues, and raise the team's DevOps maturity.
- Strong experience in production operations, technical support, and CI/CD platform engineering.
- Proven ability to troubleshoot and resolve complex application, infrastructure, and pipeline issues at the root-cause level, not just symptom-level.
- Comfortable owning both reactive support work and long-term platform/reliability improvements.
- Experience handling incidents, escalations, and production support workflows.
- Hands-on experience with IaC (Terraform, Ansible, or similar), Docker, and Kubernetes.
- Experience with CI/CD tooling (GitHub Actions, Jenkins, Azure DevOps, Harness, or similar) and pipeline standardization at scale.
- Familiarity with DevOps platform governance and security tooling (e.g., secret scanning, dependency scanning, access controls) is a strong plus.
- Comfort with automation and scripting (Python, Bash, or similar).
- Strong ownership, communication, prioritization, and problem-solving skills.
- Ability to balance reactive support work with long-term operational and platform improvements.
Tem interesse nesta vaga?