Site Reliability Engineer | Observability, Monitoring & AWS Cloud Infrastructure
Send a job offer directly to this candidate
Site Reliability / DevOps Engineer with 4+ years of experience supporting critical production systems, diagnosing complex infrastructure issues, and driving observability across Linux and AWS environments. Hands-on background in root cause analysis (RCA) and incident resolution, building monitoring, alerting, and dashboarding solutions (ELK Stack, Elastic Agents/Fleet, Beats) tied to SLO/SLA standards, and automating operational tasks with Python and Bash. Experienced with Kubernetes, Ansible-driven infrastructure automation, AWS EC2/S3, and secure access configuration (AD/LDAP, SAML SSO, TLS/SSL) and event stemming using Kafka.
Comfortable operating in on-call and cross-team support models, documenting runbooks and troubleshooting guides, and collaborating with engineering teams to improve system reliability and reduce operational overhead.
DevOps Engineer – Elastic Stack / Cloud Monitoring at Tata Consultancy Services (TCS) (2022-02 – Present)
B.Tech in Civil Engineering – Malla Reddy Institute of Technology and Science (2016 – 2020)
Intermediate (10+2) – Vagdevi Junior College (2014 – 2016)
Secondary School Certificate (SSC) – Little Buds High School (2014)