Lead Data Engineer
Smart WorkingRemotely
pythonpostgresqlsparkmongodbvector databasespgvectormilvusqdrant
Job Description
📋 Description
- Architect scalable data pipelines and infrastructure for AI/product systems.
- Design data ingestion, transformation, storage for operational/AI workloads.
- Develop batch and real-time data pipelines.
- Build vector search and ML data pipelines.
- Ensure data reliability, security and governance.
- Collaborate with AI/backend teams to support training/inference.
🎯 Requirements
- 7+ years in dedicated data engineering roles.
- Experience designing/building data pipelines and distributed systems.
- Relational DB experience; PostgreSQL preferred; MySQL acceptable.
- NoSQL experience.
- Vector databases experience.
- Strong Python programming.
🎁 Benefits
- Remote-first environment with global teams.
- Opportunity to shape data culture and hiring bar.
- Collaborate with founders and product leadership.