PuneFull TimeEngineering
Remotely
pythonsqlsparkkafkaflinkpub/subparquetavro
Job Description
📋 Description
- Pipeline Management: Maintain high-throughput streaming pipelines to ingest logs from various
- Log Normalization: Write parsers to convert raw, messy logs into standard schemas (e.g., OCSF or
- Cost Optimization: Route high-value data to the SIEM and bulk data to low-cost Object Storage (Data
- Data Preparation: Clean and structure data to enable AI/ML detection models and advanced analytics.
🎯 Requirements
- Data Engineering: Proficiency in Python (ETL) and SQL (complex querying).
- Streaming Tech: Experience with Message Queues (e.g., Kafka, Pub/Sub) and stream processing
- Log Handling: Mastery of Regex and log parsing strategies for standard formats (Syslog, CEF, JSON).
- Storage Architecture: Understanding of Data Lake principles (Parquet/Avro) vs Data Warehouses.
🎁 Benefits
- Professional development opportunities and Clearwater core values alignment.
- Competitive base salary with total rewards package including benefits and time-off policy.
Back to all jobs