Research Engineer (Reinforcement Learning)
JobgetherRemotely
pythonreinforcement learningvllmloragpussglangtrlopenrlhf
Job Description
📋 Description
- Build training environments, verifiers, and supporting infra for post-training models
- Own the synthetic data pipeline from data generation through QA
- Run end-to-end training experiments, analyze results, identify drivers
- Design and maintain evaluations for production releases
- Select/adapt open-weight foundation models for requirements
- Develop trained behaviors for voice and text agents
🎯 Requirements
- Strong Python engineering skills for production-quality systems
- Experience taking ML models from data to production
- Data-centric mindset: coverage, quality, leakage avoidance
- Ability to design robust rewards, verifiers, evaluations
- Practical experience with GPUs and their constraints
- Judgment on when training is right vs simpler approach
🎁 Benefits
- Impact on a fast-growing developer platform
- Collaboration with a small, experienced team
- Competitive salary and equity package
- Health, dental, vision benefits
- Flexible vacation policy
- Remote-friendly environment with autonomy
Back to all jobs