Job Description
📋 Description Build and evolve observability with Prometheus, Grafana, OpenTelemetry, and Elastic. Steer cloud migration and ensure reliable, performant services. Own patching and vulnerability remediation at scale (on-prem + AWS). Operate multi-tenant Kubernetes/ArgoCD and migrate core services. Contribute to AI/automation: auto-runbooks from Prometheus alerts and change-risk scoring. Consult with partner teams on metrics/thresholds and monitoring standards; mentor engineers. 🎯 Requirements 8+ years AWS experience (cloud-native and agnostic). 5+ years Kubernetes experience (EKS/AKS/GKE/Fargate). 4+ years Linux administration. Strong coding skills in Go, Python, Ruby, etc. IaC with Terraform, Ansible, CDK. CI/CD concepts and tools (Jenkins, Gradle, Maven). 🎁 Benefits Paid time off, retirement savings, equity, and stock purchase plans. Competitive health benefits and parental leave. Diverse culture with Employee Resource Groups. Equal opportunity employer; diversity and pay parity commitments.