Site Reliability Engineer in Network Infrastructure
JobgetherJob Description
📋 Description Define and manage reliability objectives for network services, including SLIs, SLOs, availability Drive reliability improvements across network infrastructure, including services, site readiness Own incident response activities, lead technical investigations, conduct postmortems, and implement Build and improve observability systems through meaningful metrics, logs, traces, alerting, and Design safer infrastructure change processes through automation, CI/CD workflows, testing Collaborate closely with network engineers and platform teams to improve system operability and 🎯 Requirements Strong knowledge of production Linux environments and a structured approach to troubleshooting Solid understanding of networking fundamentals, including control plane and data plane concepts Experience operating high-availability systems and continuously improving their reliability over Ability to write and maintain automation and infrastructure software, with experience in Go Experience with modern infrastructure tooling, including Infrastructure as Code, CI/CD systems, and Strong engineering mindset with the ability to balance operational stability, automation, and 🎁 Benefits Competitive compensation package. Career growth and continuous learning opportunities. Flexible working environment with a strong focus on ownership and autonomy. Opportunity to work on impactful cloud and AI infrastructure projects. Collaborative culture with experienced international engineering teams. Exposure to cutting-edge technologies in networking, reliability engineering, and cloud platforms.