Job Description
📋 Description Architect and optimize large-scale data platforms on Google Cloud, with BigQuery as the analytical Design and build unified batch and streaming pipelines for high-volume workloads Lead infrastructure-as-code practices for repeatable, secure, version-controlled environments Implement open table formats to enable cross-cloud/inter-engine interoperability Establish automated data quality, metadata, and lineage across the data estate Partner with data scientists, analysts, and product teams to translate business needs into reliable 🎯 Requirements 7+ years in data engineering, with at least 2 years in a lead or senior IC capacity on Google Cloud BigQuery (Advanced): deep knowledge of architecture, partitioning, clustering, slot management Streaming & Batch Pipelines: hands-on with Dataflow (Apache Beam), Dataproc, and Pub/Sub Infrastructure as Code: production experience with Terraform Open Table Formats: working knowledge of Apache Iceberg for cross-cloud interoperability Data Governance: experience with Dataplex and Data Catalog for automated quality checks, metadata 🎁 Benefits Compensation & Benefits: comprehensive health insurance, PTO, holidays, sick leave, parental Benefits details at https://egen.ai/people/#benefits