Senior Site Reliability Engineer
MegaportRemotely
gopythonawskuberneteslinuxterraformbashpostgres
Also available on
Job Description
📋 Description
- Improving production reliability and system resilience within an SRE scoped team
- Championing high standards of work and industry best practices
- Communicating with teams and stakeholders at all stages
- Bringing fresh ideas to the table and encouraging others
- Diving into complex technical problems with a can-do attitude
- Working across numerous technologies in a fast-changing industry
🎯 Requirements
- 5+ years administering Linux systems and related infrastructure in production environments
- A collaborative SRE mindset, with familiarity around SLIs/SLOs/SLAs, error budgets, blast radius
- A focus on automation, reducing toil, and preventing problem recurrence
- A track record of writing runbooks that work for the broader team, not just yourself
- Strong Kubernetes and broader ecosystem fundamentals
- Cloud infrastructure experience; AWS strongly preferred and bare-metal is a bonus
🎁 Benefits
- Flexible working environment – a remote-first culture with coworking options available.
- Generous leave plans – including 4 weeks of paid annual leave, parental leave, birthday leave, and
- Health and wellness support – through a wellness allowance and employee wellbeing initiatives.
- Comprehensive learning support – generous study and training allowance plus 5 days of paid study
- Creative, modern workspaces – designed to inspire when you're not working remotely
- Motivated, inclusive team – work alongside industry experts and fresh talent