Remotely
llmtransformersmultimodalfine tuningalignment
Job Description
📋 Description
- Design and build end-to-end agentic systems for creative tasks
- Develop novel approaches for training and adapting the large language models that power these agents
- Design new objectives, datasets, and fine-tuning strategies to improve agent behavior and
- Explore multimodal reasoning and structured generation for creative control
- Run systematic experiments to evaluate and improve agent performance in real-world tasks
- Design evaluation frameworks for agentic workflows in video analysis and editing
🎯 Requirements
- BS/MS/PhD in CS, ML, or related field
- Strong track record building production ML systems or agentic pipelines
- Deep understanding of transformers and modern LLM techniques
- Experience with fine-tuning, alignment, or post-training methods, especially for adapting models to
- Comfort owning the full stack, from model-level experiments to deployed agent systems
- Strong experimental rigor and good taste for what makes agents actually work in practice
🎁 Benefits
- Comprehensive medical, dental, and vision plans
- 401K with employer match
- Commuter Benefits
- Catered lunch multiple days per week
- Dinner stipend every night if you're working late and want a bite!
- Grubhub subscription
Back to all jobs