Also available on
Job Description
📋 Description Design and build reliability-focused frameworks, tooling, and standards that improve platform Drive initiatives that move reliability from reactive response to proactive engineering Partner with engineering teams to embed reliability into system design, development practices, and Establish and evolve observability practices, including metrics, logging, alerting, and dashboards Identify systemic risks and failure patterns, and lead efforts to address them through automation Contribute hands-on to production codebases, internal tools, and platform services with a focus on 🎯 Requirements 6+ years of experience building and operating production-grade software systems in complex Strong experience developing backend services in PHP, with the ability and desire to work Proven experience working on reliable, scalable, and highly available systems, with a clear Practical experience designing or contributing to platform-level reliability solutions, such as Solid understanding of distributed system fundamentals, including failure modes, latency, capacity Experience defining and improving operational practices, including incident response, post-incident 🎁 Benefits Remote first culture Great work-life balance with our Flexi-time policy Family Friendly policies (Enhanced Maternity and Paternity Pay and Shared Parental Leave) A dedicated company training budget Bike2Work Scheme Lifeworks, Employee Assistance Programme for wellbeing and discounts