Senior AI ML Operations Engineer
JobgetherRemotely
kubernetesterraformdatabrickslangchainlanggraphmlflowknowledge graphsvector search
Job Description
📋 Description
- Design, deploy, and maintain stable AI/ML operations platforms.
- Package and manage AI/ML services in production for reliability.
- Build automated CI/CD pipelines for model deployment.
- Provision and optimize infra for AI/ML training/inference.
- Monitor model performance, data drift, and platform health.
- Automate retraining and data workflows to sustain accuracy.
🎯 Requirements
- Proven experience with enterprise SaaS, high availability, scalability.
- Strong track record in AI/ML operations at scale.
- Hands-on with Databricks Lakehouse, MLflow, LangChain/LangGraph.
- Experience with AWS, Azure, or GCP; AWS preferred.
- Proficient in Python and SQL; Kubernetes, Docker, CI/CD, IaC.
- IaC tooling (Terraform) and observability practices.
🎁 Benefits
- Competitive compensation; enterprise-scale projects.
- Exposure to Databricks, Lakehouse, MLflow, LangChain, LangGraph.
- Opportunity to innovate and grow in AI-native engineering.
Back to all jobs