Site Reliability Engineer in Network Infrastructure
JobgetherJob Description
📋 Description Improve reliability, performance, and maturity of large-scale network systems. Define SLIs/SLOs, availability targets, and error budgets. Drive reliability improvements across network infra and inter-site connectivity. Own incident response, investigations, postmortems, and long-term fixes. Build observability with metrics, logs, traces, and alerts. Automate change processes with CI/CD, testing, and auditability. 🎯 Requirements Strong experience in site reliability engineering and infrastructure ops. Linux production environments and structured troubleshooting. Networking fundamentals: control/data planes, latency, packet loss. Experience with high-availability systems and reliability improvements. Automation/infrastructure software; Go preferred, Python welcomed. Experience with IaC, CI/CD, and container platforms. 🎁 Benefits Competitive compensation package. Career growth and continuous learning opportunities. Flexible working environment with ownership and autonomy. Opportunity to work on cloud and AI infrastructure projects. Collaborative culture with international teams. Exposure to cutting-edge networking and cloud tech.