Databricks Engineer
Blend360Remotely
pythonsqlazure databrickspysparkapache sparkdelta lakeparquetadls gen2
Job Description
📋 Description
- Design, build, and optimize scalable data pipelines on Azure Databricks
- Implement ETL/ELT workflows using PySpark, Spark SQL, and Python
- Optimize Spark jobs for performance, cost, and scalability
- Work with structured and semi-structured data (Parquet, Delta, JSON, CSV)
- Build and manage Delta Lake tables (ACID, time travel, schema evolution)
- Integrate Databricks with Azure Data Lake Storage (ADLS Gen2)
🎯 Requirements
- 4+ years of experience in Data Engineering
- Strong hands-on experience with Azure Databricks
- Proficiency in Python for data processing
- Strong knowledge of SQL (joins, window functions, performance tuning)
- Hands-on experience with Apache Spark / PySpark
- Experience with Delta Lake
🎁 Benefits
- Experience with Azure Data Factory
- Exposure to CI/CD pipelines (Azure DevOps, GitHub Actions)
- Basic understanding of data modeling
- Familiarity with cloud security and RBAC in Azure
- Exposure to streaming data (Spark Structured Streaming, Event Hub, Kafka)