Remotely
pythoncomputer visionbiometricsreal time systemspytorchmultimodal models
Job Description
📋 Description
- Research and develop multimodal perception and authentication methods across visual, audio, and
- Explore how specialized perception models and larger multimodal models can work together.
- Design data, training, and evaluation approaches that improve performance in real-world conditions.
- Study model behavior, robustness, and failure modes across sensing, data, and deployment
- Integrate and validate new capabilities in real-time or resource-constrained systems.
- Work with hardware, firmware, software, and product teams to turn research into working systems.
🎯 Requirements
- Strong background in computer vision, audio or speech ML, multimodal learning, or sensing.
- Experience developing specialized ML models, larger multimodal models, or both.
- Proven ability to bring research ideas into practical systems, prototypes, or products.
- Skill in designing experiments, building evaluations, and investigating model behavior.
- Experience with sensing hardware, real-time systems, or deployment constraints.
- Proficiency in Python and PyTorch; comfortable with C++ or systems integration.
🎁 Benefits
- Relocation assistance for eligible candidates.
- Hybrid work model with in-office presence three days per week.
- OpenAI equity options as part of compensation package.