Site Reliability Engineer
HelsingJob Description
📋 Description Design, implement, and manage on‑premise Kubernetes infrastructure. Build cloud‑native platforms on‑premises for scalable services. Establish observability with Grafana, Prometheus, and distributed tracing. Architect secure, multi‑tenant clusters with policy‑as‑code and zero‑trust networking. Develop and maintain MLOps platforms to deploy and monitor ML models. Collaborate with Security teams on supply chain security and runtime protection. 🎯 Requirements Scripting: Python, Go, Rust or Bash/Shell for automation. Experience with GitOps workflows and CI/CD automation. Kubernetes production experience, custom controllers/operators, service mesh (Istio/Linkerd). Cloud‑Native Tech: Helm, ArgoCD, Flux; container security tools like Falco. Observability: Grafana, Prometheus, Loki, Tempo, OpenTelemetry; dashboards and SLIs/SLAs. Networking: strong networking concepts and security. 🎁 Benefits Focus on outcomes, not time‑tracking. Competitive compensation and VSOP options. Relocation support. Social and education allowances. Regular company events and all‑hands across Europe. Infraduction onboarding to learn our stack and tools from day one.