Senior Site Reliability Engineer I
BrazeTorontoFull TimeEngineering
Remotely
gokuberneteslinuxansibleterraformrubynginxslo
Job Description
📋 Description
- Lead NGINX & Kubernetes Ingress Infrastructure
- Architect, operate and tune Braze's high-performance ingress layers for massive real-time API
- Own scaling routines using RED metrics, HPA, and custom policies for high-throughput services
- Partner with product teams to translate requirements into resilient, scalable stacks
- Manage SLIs, SLOs, and error budgets for API services
- Conduct systems design and capacity planning to meet enterprise SLAs
🎯 Requirements
- 5+ years in DevOps or SRE in a high-scale production environment
- Deep hands-on NGINX configuration, troubleshooting, and ingress management
- Proficiency with Kubernetes administration and container orchestration
- Strong Linux/Unix fundamentals
- Programming/scripting in Ruby and/or Go (or Python/Java) for tooling
- Infrastructure as Code experience (Terraform, Ansible, Chef, etc.)
🎁 Benefits
- Comprehensive benefits varies by location
- Competitive compensation, potential equity
- Flexible PTO and total rewards package
- Healthcare benefits including medical, dental, vision
- Professional development support and learning stipend
- Inclusive, collaborative culture with Great Place to Work recognition
Back to all jobs