Job Description
📋 Description Lead and grow a high-performing team of 6 engineers. Architect scalable ML runtime systems for edge and cloud. Balance real-time latency with high-throughput cloud serving. Transition ML workloads to a JAX-native runtime (OpenXLA/PjRT, TensorRT). Collaborate with ML researchers to optimize workloads. Develop profiling/benchmarking to remove bottlenecks. 🎯 Requirements B.S. or M.S. in CS, EE, Deep Learning, or related field. People mgmt experience with senior engineers. 8+ years in software eng for ML systems and infra. Strong production programming expertise. Proven track record optimizing ML for hardware (GPUs/TPUs). Hands-on in distributed backend systems, low-latency & fault-tolerant. 🎁 Benefits Discretionary annual bonus program. Equity incentive plan. Generous company benefits program.