San Francisco, San JoseFull TimeEngineering
Remotely
kuberneteslinuxnetworkingcloud platformshpcstoragegpu
Job Description
📋 Description
- Take signed deployments from contract to live production
- Validate configuration, connectivity, storage, compute
- Collaborate with engineering to resolve gaps
- Maintain accurate technical deployment pictures
- Provide regular status updates to stakeholders
- Guide onboarding to first production workload
🎯 Requirements
- 4+ years of hands-on technical experience with GPU/HPC infra, cloud, Kubernetes, or large-scale
- Strong troubleshooting and ability to dive into networking/storage/compute
- Coordinate across eng and infra teams to close dependencies
- Clear written and verbal communication for status updates
- Own deployment details and risk factors
- Bias for ownership and action-oriented work
🎁 Benefits
- Generous cash & equity compensation
- Health, dental, and vision coverage
- Wellness and commuter stipends
- 401k with 2% company match (USA)
- Flexible paid time off