Description :
Senior Platform Engineer Bengaluru (610 Years)
Role Overview :
You will design, build, and scale core platform systems that empower engineers to develop, deploy, and operate services efficiently. This role emphasizes Python-based automation, cloud-native infrastructure, and Kubernetes reliability engineering.
Key Responsibilities :
Reliability & Operations :
- Design, build, and maintain scalable platform services and developer tooling using Python.
- Develop automation frameworks, orchestration systems, and backend infrastructure services.
- Improve platform reliability, observability, deployment automation, and operational excellence.
- Implement best practices around security, monitoring, testing, and production operations.
- Lead incident response, RCA, and postmortems; drive reliability improvements through automation.
- Contribute to architectural decisions for distributed systems and cloud-native applications.
- Mentor engineers and promote engineering excellence across the organization.
Cloud & Platform Engineering :
- Build and manage infrastructure on AWS.
- Operate Kubernetes clusters (EKS preferred).
- Deploy services using Helm, ArgoCD, Argo Rollout.
- Manage containerized workloads using Docker / containerd.
- Implement Infrastructure-as-Code (Terraform).
- Automate with Python, Ansible, Packer.
- Integrate CI/CD pipelines with GitHub Actions / Jenkins / ArgoCD.
- Build observability solutions with Prometheus, Grafana, Datadog, Splunk.
Automation & Tooling :
- Develop Python-based automation and reliability tooling.
- Create internal monitoring and operational tools.
- Integrate CI/CD pipelines with observability and reliability checks.
Collaboration & Leadership :
- Mentor junior engineers and influence architecture decisions.
- Collaborate across engineering, product, and security teams.
Required Qualifications :
- 610 years of experience in SRE, DevOps, or Platform Engineering.
- Strong Python programming skills for production-grade automation.
- Hands-on Kubernetes experience (EKS preferred).
- Solid understanding of observability fundamentals.
- Experience with Helm, ArgoCD, Argo Rollout, Docker, Ansible, Packer.
- Strong AWS cloud experience.
- Solid Linux and networking fundamentals.
- Familiarity with SDLC and modern DevOps practices.
Preferred Qualifications :
- Toolchain development using Python.
- Multi-cluster or multi-region Kubernetes experience.
- Service mesh (Istio) and API gateway (Kong) exposure.
- Cloud cost optimization expertise.
- Advanced Infrastructure-as-Code with Terraform.
Project Focus :
- Build internal platform tools and APIs using Python.
- Automate infrastructure provisioning, deployments, and operational workflows.
- Design scalable backend services and platform components.
- Improve CI/CD pipelines, deployment reliability, and developer experience.
- Manage and optimize Kubernetes and AWS infrastructure.
- Build observability solutions (monitoring, logging, alerting, tracing).
- Reduce operational toil through automation and self-service platforms.
- Partner with engineering and security teams to drive scalability and reliability.