Job Description
📋 Description Design and optimize real-time speech-to-text pipelines (ASR, VAD) Improve transcription accuracy via context injection (names, teams, vocab) Develop and maintain LLM-powered post-processing (grammar, filler removal) Build voice-to-action systems that parse natural language into workspace commands Evaluate and integrate ASR models (Whisper, AssemblyAI, Fireworks) for cost and latency Explore multimodal AI (screen/voice/text) for next-gen assistant experiences 🎯 Requirements Real-time speech-to-text pipeline design and optimization Experience with ASR models and audio processing Experience with LLMs and NLP tasks Proficiency in benchmarking ML models and evaluating cost/latency Design voice-driven actions and command parsing Cross-functional collaboration with product/platform teams 🎁 Benefits Equal Opportunity Employer Privacy Notice Visa sponsorship for eligible engineering/product roles (not guaranteed)