Site Reliability Engineer, Litmus GCC
LitmusJob Description
📋 Description Own day-to-day reliability, security, and performance of Azure-hosted UNS and LEM environments for Provision, configure, and maintain cloud infrastructure spanning compute, container orchestration Monitor, alerting, dashboards, and incident response to meet 99.9% uptime commitments. Manage identity/SSO integration, network security, backups, DR, and capacity planning. Implement infrastructure-as-code and CI/CD pipelines to automate provisioning and deployment. Collaborate with customer stakeholders during validation, go-live, and support activities. 🎯 Requirements 4-8 years in SRE/DevOps/Cloud Infra with strong Azure experience. Hands-on with Azure services: AKS, VMs, VNet/NSG/_Load Balancer/App Gateway, Azure DB for Production Kubernetes administration and troubleshooting. IaC tools: Terraform, Bicep, or ARM templates. Scripting: Python, Bash, PowerShell. Networking fundamentals: DNS, TLS/SSL, load balancing, firewalls/NSGs.