Job Description
📋 Description Hybrid role building the onboard ML inference engine for Waymo's models across perception Work across the ML stack from efficient models, model compression, and ML software (e.g., JAX, XLA Collaborate with ML practitioners and hardware teams to optimize models for NVIDIA GPUs and custom 🎯 Requirements B.S. or M.S. in CS, EE, Deep Learning or related field 5+ years of industry experience on system performance, hardware-level GPU optimization, or ML Strong C++ and CUDA programming skills Extensive experience in NVIDIA GPU Kernel development to accelerate deep learning models Proven debugging and optimization experience on the XLA:GPU compiler and NVIDIA runtime stack Passion for developing and optimizing ML software stacks for modern ML accelerator architectures 🎁 Benefits Salary range: $213,000 – $263,000 USD Discretionary annual bonus program and equity incentive plan Generous company benefits program