Zscaler
About this role
Zscaler runs a cloud security platform that inspects a very large volume of enterprise traffic, so its SRE work is genuinely large scale. This seat is on the SRE Cloud Infrastructure and Operations team, reporting to the Director of Site Reliability Engineering, and the posting is clear that the job is architecting and automating rather than only responding: designing advanced cloud management automation to eliminate toil, owning cloud operations, deployments, on call support and incident management, and tuning Linux and BSD based systems. The technical list is deep and unusually broad, spanning C, Java, Go and Python, Terraform and Ansible, Kubernetes and AWS, plus observability work in Grafana with SLIs, SLOs and error budgets. Note that BSD experience is named alongside Linux, which is rarer and worth highlighting if you have it.
Who this is for
Required by the posting: a minimum of 7 years of relevant experience designing, analysing and troubleshooting large scale distributed systems. Deep hands on experience with C, Java, GoLang, Python, Terraform, Ansible, Python automation, networking, Kubernetes and AWS cloud. Proven experience in observability, building complex dashboards, managing Grafana, and a sharp understanding of SLIs, SLOs and error budgets. Modern DevOps expertise. Web protocol depth. Demonstrated curiosity and active exploration of AI tools, with a history of integrating new technology into daily workflows, which Zscaler states as a requirement rather than a preference across its postings.
The day to day: design, implement and manage advanced cloud management automation to eliminate toil and accelerate delivery. Own cloud operations, deployments, on call support and incident management. Continuously design and tune Linux and BSD based systems. The role reports directly to the Director of Site Reliability Engineering on the SRE Cloud Infrastructure and Operations team.
Location: Hyderabad. Division is Engineering, region India, employment type full time employee.
Who should apply: SREs seven or more years in with real distributed systems debugging depth and at least one systems language in their hands. The on call responsibility is stated in the posting rather than buried, so weigh it. If you have FreeBSD or other BSD operational experience, this is one of the few postings where it is an explicit advantage rather than trivia.
The day to day: design, implement and manage advanced cloud management automation to eliminate toil and accelerate delivery. Own cloud operations, deployments, on call support and incident management. Continuously design and tune Linux and BSD based systems. The role reports directly to the Director of Site Reliability Engineering on the SRE Cloud Infrastructure and Operations team.
Location: Hyderabad. Division is Engineering, region India, employment type full time employee.
Who should apply: SREs seven or more years in with real distributed systems debugging depth and at least one systems language in their hands. The on call responsibility is stated in the posting rather than buried, so weigh it. If you have FreeBSD or other BSD operational experience, this is one of the few postings where it is an explicit advantage rather than trivia.
Apply on company site
Opens job-boards.greenhouse.io, the employer's own application page. Applying is always free.