Senior Technical Operations & Deployment Engineer (GPU Cloud Infrastructure)
JobgetherRemotely
kuberneteslinuxansibleterraformprometheusgrafanacudanvidia
Job Description
📋 Description
- Hands-on role deploying, commissioning, and operating GPU cloud environments
- Turn validated architectures and BOMs into production-ready infrastructure
- Work across hardware, networking, storage, Linux, and platform software
- Collaborate with datacenter ops, network engineering, and cloud platform teams
- Troubleshoot complex issues across physical and software layers
- Establish deployment standards, validation procedures, and runbooks
🎯 Requirements
- Datacenter infrastructure experience in GPU/HPC/AI cloud environments
- Bare-metal deployment from physical install to production readiness
- Experience with NVIDIA GPU servers, drivers, RAID/PCIe, high-density compute
- Familiarity with NVLink/NVSwitch, RoCE/RDMA, in-rack networking
- Linux troubleshooting and admin, virtualization and containers (KVM/QEMU, Docker, Kubernetes)
- Automation and scripting (Terraform, Ansible, Bash, Python)
🎁 Benefits
- Attractive compensation package
- Full-time or contract engagement
- Europe-based remote working environment
- Hands-on exposure to cutting-edge GPU cloud and AI infra
- Cross-functional collaboration across hardware, networking, storage, Linux, cloud
Back to all jobs