Datacentre Operations Engineer
RadiantRemotely
aihpccudagpunvidiadata centregpu compute
Job Description
📋 Description
- On-site hardware operations for East London AI infrastructure site
- Hands-on deployment, maintenance, and break/fix for high-density, air-cooled compute
- Respond to hardware alerts and support rapid incident response
- Collaborate with HPC SRE, Network Engineering, and Datacentre Strategy teams
- Uphold world-class SLAs across multi-data hall environments
🎯 Requirements
- Degree in CS/EE or 5+ years data centre ops experience
- 3+ years in data centre operations, HPC, or related roles
- Proven hands-on experience with HPC/NVIDIA GPU platforms
- Knowledge of high-density air cooling, CRAC/CRAH, containment, and throughput monitoring
- Strong English communication; ability to work 24x7 with on-site shifts
- Willingness to travel within EMEA and upskill in SRE practices
🎁 Benefits
- Exposure to world-class GPU/AI infrastructure
- Global, inclusive engineering culture with growth opportunities
- Structured on-site shift pattern ensuring SLA delivery
Back to all jobs