Senior Site Reliability Engineer
SanityJob Description
📋 Description Design, build, and operate GCP infra, Kubernetes, networking, CI/CD, and observability. Diagnose and troubleshoot complex distributed systems at high request volume. Ensure observability and analyze behavior of our stack. Modernize edge, caching, and gateway layers with Fastly and improved observability. Raise reliability with dashboards, alerting, paging standards, and on-call readiness. Make deployments boring: golden paths, readiness checks, safe rollouts, and automation. 🎯 Requirements Based in the United States with overlap with European hours. 5+ years of SRE/on-call experience. Kubernetes for container orchestration in cloud environments. CI/CD pipeline experience. Observability stack expertise (Prometheus, etc.). Comfortable with incidents and high-pressure scenarios. 🎁 Benefits Highly skilled, inspiring, and supportive team. Real infrastructure scale and meaningful, hands-on work. Flexible, trust-based environment with growth opportunities. Global, diverse group of colleagues and customers. Comprehensive health plans and perks. Healthy work-life balance for individuals and families.