ClickHouse
About this role
This is the reliability side of the same ClickHouse Postgres effort. The company is expanding its cloud platform across AWS, GCP and Azure, and it wants an SRE who owns how those services run: upgrades, patching, maintenance and scaling, plus the provisioning and deployment automation underneath. The posting is explicit that this is hands on rather than advisory, so you write Go tooling, build Terraform and CI/CD pipelines, own observability with tools like Prometheus, Grafana, Loki and OpenTelemetry, and run incident management and postmortems. Apply if you have run production distributed systems for several years, know Postgres operations and tuning, and are comfortable in a multi cloud topology rather than one provider. It is remote in India, which makes it a rare remote infrastructure job rather than another Bangalore hybrid seat.
Who this is for
Required by the posting: 7+ years in SRE, DevOps or infrastructure engineering with a track record of running distributed, production grade systems. Solid Postgres operations, scaling and performance tuning. Deep hands on AWS with exposure to GCP and Azure, and comfort navigating multi cloud topologies. Proficiency with Terraform, Kubernetes and container based infrastructure. Strong Go development skills, or a stated willingness to write and own production Go. Familiarity with observability tooling such as Prometheus, Grafana, Loki or OpenTelemetry. Deep understanding of SLOs, incident response and continuous reliability improvement.
The real day to day: leading reliability and operations for the Postgres integration including upgrades, patching, maintenance and scaling; designing provisioning, deployment and lifecycle automation across the three major clouds; writing infrastructure as code; contributing Go tooling; owning alerting, metrics and tracing; driving incident management and postmortem practice; and mentoring engineers as the team scales.
Location and working pattern: India (remote). ClickHouse operates in over 20 countries and provides a home office setup allowance to remote employees. The same requisition also exists for the United States; this is the India one.
Honest fit guidance: the posting asks for a "founder's mentality" and describes the role as hands on and high impact, which in practice means the automation, the code and the on call are all yours rather than split across a platform team. If you are an SRE who mostly operates other people's tooling and does not write production Go, expect that to be the gap they probe. Postgres operations depth is the other hard requirement and is harder to fake than the cloud list.
The real day to day: leading reliability and operations for the Postgres integration including upgrades, patching, maintenance and scaling; designing provisioning, deployment and lifecycle automation across the three major clouds; writing infrastructure as code; contributing Go tooling; owning alerting, metrics and tracing; driving incident management and postmortem practice; and mentoring engineers as the team scales.
Location and working pattern: India (remote). ClickHouse operates in over 20 countries and provides a home office setup allowance to remote employees. The same requisition also exists for the United States; this is the India one.
Honest fit guidance: the posting asks for a "founder's mentality" and describes the role as hands on and high impact, which in practice means the automation, the code and the on call are all yours rather than split across a platform team. If you are an SRE who mostly operates other people's tooling and does not write production Go, expect that to be the gap they probe. Postgres operations depth is the other hard requirement and is harder to fake than the cloud list.
Apply on company site
Opens job-boards.greenhouse.io, the employer's own application page. Applying is always free.