Sr. Site Reliability Engineer (Azure, IaC, Distributed Systems)
JobgetherRemotely
awsinfrastructure as codeazuregcpansibleterraformiaccloudformation
Job Description
📋 Description
- Support deployment, configuration, and maintenance of monitoring/logging across envs.
- Maintain observability tech: Splunk, Grafana, Prometheus, OpenTelemetry, CloudWatch, etc.
- Automate operational activities to improve workflows and integration.
- Contribute to CI/CD pipelines for automated build, test, deploy.
- Manage cloud infra on AWS and GCP for availability and security.
- Use IaC tools Terraform/Ansible/CloudFormation to configure envs.
🎯 Requirements
- Bachelor’s degree + 2+ yrs experience or 5+ yrs exp without degree.
- Hands-on with monitoring/logging deployments in prod.
- Experience with observability platforms: Splunk, Grafana, Prometheus, OpenTelemetry, Fluent Bit
- Cloud infra experience AWS/GCP; knowledge of availability and security.
- IAM/Networking basics; IaC: Terraform/Ansible/CloudFormation.
- CI/CD pipelines knowledge; Docker & Kubernetes familiarity.
🎁 Benefits
- Fully remote work model in India.
- Full-time with Mon-Fri schedule.
- Exposure to modern cloud, observability, automation tech.
- Work with large-scale tech and distributed systems.
- Opportunities to grow in DevOps/SRE and cloud infra.
- Collaborative environment across global teams.