Remotely
awskubernetesredisobservabilityaiopensearchclauderds
Job Description
📋 Description
- Extend the self-service datastore platform provisioning automation, guardrails, and paved paths for
- Ship observability, alerting, and backup/disaster recovery as built-in defaults with every datastore
- Turn recurring reliability issues, scaling questions, and schema-change safety into platform
- Codify compliance, capacity management, and cost management into the platform
- Support the operational needs of the growing database fleet alongside product teams
- Other projects as business needs require
🎯 Requirements
- 3+ years of experience as a Site Reliability Engineer or similar
- Proven track record building/maintaining production infrastructure in AWS and Kubernetes
- Experience delivering tooling, platforms, or production services to users
- Strong fluency with AI coding tools like Claude
- Adaptable with unique skills and perspectives; encouraged to apply even if not 100% aligned
🎁 Benefits
Back to all jobs