Senior Technical Operations & Deployment Engineer (GPU Cloud Infrastructure)
JobgetherRemotely
dockerkuberneteslinuxansibleterraformnvmegpunvidia
Job Description
📋 Description
- Highly hands-on role focused on deploying, commissioning, and operating GPU cloud environments
- Turn validated architectures and BOMs into production-ready infrastructure spanning hardware
- Work at the intersection of datacenter operations, GPU infrastructure, network engineering, and
- Operate with NVIDIA GPU systems, advanced networking, storage platforms, Kubernetes
- Troubleshoot complex issues across physical and software layers; drive incidents through resolution.
- Help establish deployment standards, validation procedures, documentation, and operational
🎯 Requirements
- Datacenter infrastructure experience, ideally in GPU, HPC, AI cloud, private cloud, or high-density
- Bare-metal deployment from physical install to validation and production readiness.
- GPU infrastructure experience with NVIDIA GPU servers, drivers, and hardware validation.
- Familiarity with NVL72-style rack-scale architectures, NVLink/NVSwitch domains, in-rack networking
- Linux troubleshooting across OS, kernels, drivers, and hardware interfaces.
- Networking knowledge: VLANs, VRFs, BGP, ECMP, OVS/OVN, high-speed datacenter connectivity.
🎁 Benefits
- Europe-based remote working environment with flexibility.
- Hands-on exposure to NVIDIA GPU platforms, RoCE/RDMA networking, Kubernetes, virtualization
- Broad cross-functional scope across hardware, datacenter operations, networking, storage, Linux
- High-impact role within a fast-growing international scale-up.
- Strong opportunities for technical growth as infrastructure expands.
- Friendly, diverse, and international working environment.
Back to all jobs