Palo AltoFull TimeEngineering
Remotely
pythonsqlsparkkafkaflinkhadoopstatistical methods
Job Description
📋 Description
- Build and maintain data pipelines for high-volume datasets (Spark, Kafka, Flink)
- Develop quantitative models and statistical frameworks for experimentation and forecasting
- Create data infrastructure ensuring data quality and accessibility
- Collaborate with product engineering, product, and operations to production-grade systems
- Conduct A/B tests, causal analysis, and performance evaluations
- Implement monitoring and automation for real-time decision support
🎯 Requirements
- 4+ years building production data pipelines at scale
- Proficient in Python, SQL, and distributed computing (Spark, Flink, Hadoop)
- Expertise in statistical methods, predictive modeling, experimental design
- Cloud services for data storage, processing, and orchestration
- Bachelor's or Master's in CS, Statistics, Applied Math, or related field
- Excellent problem-solving with focus on business impact through reliable systems
🎁 Benefits
- $180,000 - $440,000 USD
- Equity, medical, vision, dental, 401(k), disability, life insurance, discounts and perks