Mountain ViewFull TimeEngineering
Remotely
machine learningdistributed systemsllmshigh throughputlow latencyinferencemodel hosting
Job Description
📋 Description Develop Waymo’s inference platform to make it scalable, high throughput, and low latency Host internal and external ML models, including LLMs, with other teams across Waymo Improve the efficiency of running inference on large models to increase throughput and save cost Deploy and integrate model inference solutions across use cases like distillation, eval, dataset 🎯 Requirements 5+ years of professional experience in software engineering Experience in programming C++ Experience with building highly scalable distributed systems 🎁 Benefits Waymo discretionary annual bonus program Equity incentive plan Generous company benefits program
Back to all jobs