San FranciscoFull TimeEngineering
Remotely
pythonpytorchjaxcudatritondistributed trainingmodel parallelism
Job Description
📋 Description
- Member of Technical Staff specializing in ML/LLMs.
- Early-stage AI company building foundation models.
- Work across LLM research, training infra, post-training, GPU optimization.
- Collaborate on ownership across the stack in a small, technical team.
- Hands-on research and engineering roles with impact across training.
- On-site in San Francisco.
🎯 Requirements
- 1+ years in theoretical LLM research or ML engineer role.
- Hands-on language model experience beyond APIs.
- Experience with LLM architecture, pre/post-training, RL or training infra.
- Strong Python, PyTorch, CUDA, Triton experience.
- Distributed training, model/data parallelism, GPU optimization.
🎁 Benefits
- Base salary $200,000 – $350,000 USD.
- Equity compensation.
Back to all jobs