Remotely
pythonsqldockerkubernetessparkbigqueryairflowhadoop
Also available on
Job Description
📋 Description
- Design and implement ETL pipelines to transform upstream data into curated assets.
- Architect highly distributed, high-performance data platform systems.
- Identify and improve existing pipelines and processes.
- Translate business requirements into technical solutions.
- Ship ETL improvements and fix data issues in the data stack within the first 90 days.
- Collaborate with Product and Engineering teams to build data sets and answer key questions.
🎯 Requirements
- Strong SQL experience.
- Expertise with data pipelining orchestration frameworks (Airflow, Luigi, etc.).
- Experience building ETL pipelines.
- Software engineering in Java or Python.
- Experience building/optimizing a data warehouse on a major cloud platform (BigQuery preferred).
- Deep understanding of Big Data tech like Hadoop and Spark.
🎁 Benefits
- Competitive salary and stock options.
- Flexible vacation / Unlimited PTO (within reason).
- Medical, dental, and Flexible Spending Account (FSA).
- Company-matched 401(k).
- Paid Paternity & Maternity Leave.
- Agile Development Program for continued learning and professional development.