Lead ML/AI Platform Engineer
JobgetherRemotely
pythonllmgenaipytorchtensorflowsagemakermlflowbedrock
Job Description
📋 Description
- Own the ML/AI platform: lead the architecture and operation of training infrastructure, model
- Set technical direction: define ML/AI architecture, tooling standards, and build-versus-buy
- Build production ML systems: oversee feature engineering, model training, tuning, deployment
- Drive GenAI and LLM engineering: strategies for RAG, prompt engineering, evaluation, fine-tuning
- Develop agentic capabilities: evaluate emerging agent technologies and guardrails for regulated
- Own the serving layer: design scalable ML services and API contracts with Java microservices
🎯 Requirements
- Senior engineering experience: 8+ years in software or ML, 5+ years delivering production ML
- Technical leadership: demonstrated ability to shape ML strategy and mentor engineers.
- AWS ML expertise: hands-on with SageMaker, Bedrock, AgentCore for GenAI/agentive apps.
- Open-source ML tooling: experience with JupyterLab, Spark, MLflow, and related tools.
- AWS platform knowledge: strong SQL and services like S3, Athena, Redshift, Glue, Step Functions
- GenAI/LLM expertise: production experience with RAG, prompt engineering, evaluation, trade-offs.
🎁 Benefits
- Equity compensation package.
- Flexible Time Off (FTO).
- Medical, dental, and vision coverage with premiums covered where applicable.
- Disability and life insurance.
- Learning and career development opportunities.
- Remote-work setup reimbursement.
Back to all jobs