North AmericaFull TimeEngineering
Remotely
machine learningevaluationpost trainingfine tuning
Job Description
📋 Description
- Design and run post-training workflows that improve AI system behavior
- Develop datasets, preference signals, evaluation suites, reward models, and fine-tuning workflows
- Investigate how post-training techniques affect enterprise workflows and production constraints
- Build infrastructure for experimentation, model comparison, and regression testing
- Partner with AI Researchers/Engineers to apply methods in deployed systems
- Analyze model outputs and production traces to identify improvement opportunities
🎯 Requirements
- Experience improving model behavior with fine-tuning, preference optimization, reinforcement
- Strong programming and experimentation skills; ability to build training/evaluation pipelines
- Research-oriented builder with focus on understanding why behavior changes
- AI systems mindset: data, prompts, tools, retrieval, evaluators, deployment context
- AI-native working style with daily use of AI tools for coding, analysis, and debugging
- Ownership mentality and bias towards measurement; comfort with practical constraints
🎁 Benefits
- Base salary range $150K – $250K with equity and comprehensive benefits
- 100% medical, dental, and vision insurance for employee and dependents
- Flexible time off and retirement planning (401(k), HSA/FSA)
- Wellness benefits and family-building resources
- Complimentary in-office lunches and snacks
- Access to state-of-the-art AI models and tools
Back to all jobs