Senior Site Reliability Engineer (SRE) | Feeld
GTRemotely
awsreact nativetypescriptnode.jscloudflarecloudwatchsentry
Job Description
📋 Description
- Own observability for critical product and user journeys within your squad.
- Define, build and maintain metrics, dashboards and alerts.
- Define and maintain SLIs/SLOs for key services and product-level metrics.
- Improve monitoring, logging, tracing and alerting across the squad’s systems.
- Act as first responder for critical production incidents, including out-of-hours incidents.
- Investigate production signals and begin mitigating issues independently.
🎯 Requirements
- Strong backend/software engineering background with senior-level Node.js and TypeScript experience.
- Hands-on experience in an SRE, Production Engineering or reliability-focused role.
- Production experience with AWS.
- Experience with Cloudflare and CloudWatch.
- Experience with monitoring and observability across metrics, logging, tracing and alerting.
- Experience handling production incidents, including triage, mitigation and postmortems.
🎁 Benefits
- Health insurance.
- Wellbeing budget.
- Sport coverage.
- Learning budget.
- 18 paid vacation days per year.
- Paid sick leaves.