North AmericaFull TimeEngineering
Remotely
awsdatadogdevopskubernetesterraformsredatabrickscloudflare
Job Description
📋 Description
- Design, build, and operate shared cloud infrastructure using AWS, Kubernetes, Terraform
- Deliver SRE and DevOps initiatives that improve reliability, scalability, observability, deployment
- Build reusable infrastructure modules, automation, and self-service workflows that reduce manual
- Help define and implement service-level indicators, service-level objectives, monitoring, alerting
- Participate in incident response and improve operational outcomes through clear runbooks, effective
- Strengthen disaster-recovery readiness through recovery planning, automation, testing, and
🎯 Requirements
- 4+ years of relevant software, infrastructure, SRE, platform engineering, or DevOps experience
- Strong interest in building AI-native and data-intensive applications and systems, working with
- Proficient in writing reliable, maintainable software and automation; comfortable across
- Experience operating production systems with observability, incident response, operational
- Experience improving CI/CD, infrastructure as code, deployment workflows, or internal developer
- Understanding of reliability concepts (SLIs, SLOs, error budgets, capacity planning, DR
🎁 Benefits
- We offer flexible work hours, generous benefits, 401K match, parental leave, team events, wellness
- Growth is based on impact and opportunity within the team; ownership and trust are emphasized.
- The annual base salary range is $168,000 - 193,725; equity included.
Back to all jobs