Senior Technical Operations & Deployment Engineer (GPU Cloud Infrastructure)
JobgetherRemotely
kuberneteslinuxansibleterraformrocegpunvidiardma
Job Description
📋 Description
- Senior, hands-on infrastructure role focused on deploying, commissioning, and operating GPU cloud
- Turn validated architectures and bills of materials into production-ready infrastructure spanning
- Operate at the intersection of datacenter, GPU infra, networking, and cloud platform operations.
- Work with high-density NVIDIA GPU systems, advanced networking, storage platforms, Kubernetes
- Troubleshoot complex issues across physical and software layers and drive incidents to resolution.
- Help establish deployment standards, validation procedures, documentation, and operational
🎯 Requirements
- Datacenter infrastructure deployment and maintenance experience, incl. GPU/HPC/private cloud
- Bare-metal deployment from physical install to production readiness.
- GPU infrastructure experience with NVIDIA GPUs, drivers, PCIe, and high-performance compute.
- Familiarity with NVL72 rack-scale architectures, NVLink/NVSwitch, in-rack networking, and
- Linux troubleshooting, kernel/driver/hardware interface management.
- Networking and GPU networking know-how (VLANs/VRFs, BGP/ECMP, OVS/OVN, RoCE/RDMA, SR-IOV).
🎁 Benefits
- Europe-based remote working environment with flexibility.
- Hands-on exposure to NVIDIA GPU platforms, RoCE/RDMA networking, Kubernetes, virtualization
- High-impact role in a fast-growing international scale-up.
- Opportunities for technical growth as the infra expands.
- Friendly, diverse, and international working environment.
- Inclusive workplace committed to equal opportunity.
Back to all jobs