Staff Site Reliability Engineer
GradleAlso available on
Job Description
📋 Description Operate and maintain Develocity production instances Lead SRE practices: on-call, incident response, SLOs Troubleshoot across stack with our Cloud Platform Build reliability into systems through cross-team collaboration Mentor SREs and onboard new teammates 🎯 Requirements 7+ years in SRE/DevOps at scale Experience leading reliability initiatives across teams Design and operate systems with SLOs and error budgets Strong Kubernetes experience in production Cloud infrastructure expertise, preferably AWS Python and Bash scripting for automation 🎁 Benefits Ground-floor role shaping SRE practices Ownership of production systems used by engineers Direct customer interaction during incidents Automation-focused culture over heroics Remote-friendly with offsite team meetings Competitive salaries and equity grants