Loading…
Loading…
<p>Ensure the reliability, performance, and scalability of our automation platform. You'll own our Kubernetes clusters, observability stack, alerting pipelines, and incident response runbooks for a 24/7 enterprise product.</p><h3>Requirements</h3><ul><li>4+ years of SRE or DevOps experience</li><li>Expert-level Kubernetes administration</li><li>Experience with observability stacks (Prometheus, Grafana, Loki)</li><li>Strong on-call and incident management experience</li><li>Infrastructure-as-code experience (Terraform, Pulumi)</li></ul>