Site Reliability Engineer
NiceRemotely
gopythonawsdockerkubernetesterraformnodeprometheus
Job Description
📋 Description
- Run production systems by monitoring availability and health
- Build software/systems to manage platform infrastructure
- Improve reliability, quality, and time-to-market
- Analyze metrics to tune performance and fault finding
- Provide operational support for large distributed apps
- Collaborate with teams on design, testing, and capacity planning
🎯 Requirements
- 4+ years programming/scripting (Go, Python, C#, Node)
- Bachelor's in CS/Engineering or equivalent
- 6-8 years in similar role focusing on systems engineering and automation
- Proficient in at least one language (Python/Go/Java/C#) and scripting (Bash/PowerShell)
- Experience with AWS and cloud services (EC2/ECS/Lambda)
- Infrastructure as code (CloudFormation, Terraform)
🎁 Benefits
- NiCE-FLEX hybrid model: 2 days in-office, 3 days remote
- Global, growth-focused environment with internal career opportunities
- Collaborative, innovative team culture
Back to all jobs