Head of SRE
JobgetherRemotely
awsdevopskubernetesazureterraformsresite reliabilityci/cd
Job Description
📋 Description
- Own SRE strategy, roadmap, and execution across teams.
- Assess infra, modernize or rebuild where needed.
- Architect and run scalable, secure prod environments in cloud.
- Define SLOs, uptime targets, reliability standards.
- Establish monitoring, logging, observability for prods.
- Design incident management, RCAs, escalation, postmortems.
🎯 Requirements
- Proven hands-on SRE / Production Engineering / DevOps experience.
- Cloud expertise with AWS or Azure; IaC (Terraform); Kubernetes.
- Experience maturing SRE practices at scale.
- Strong uptime, reliability, performance track record.
- CI/CD, observability, automation, modern delivery.
- Incidence response, on-call, and ops readiness frameworks.
🎁 Benefits
- Fully remote within the European time zone.
- Full-time role with significant ownership across engineering.
- Lead SRE function and reliability standards organization-wide.
- Direct exposure to product and engineering leadership.
- Opportunity to work on large-scale AI infra and ML systems.
- Autonomy to improve and rebuild infra and processes.
Back to all jobs