Also available on
Job Description
📋 Description Build LLM observability frameworks, libraries, pipelines, and APIs. Collaborate with the AI OSS community; gather feedback and review PRs. Prototype and iterate rapidly on state-of-the-art LLM techniques. Improve observability and debugging; surface insights on LLM behavior to diagnose issues. Educate and evangelize with blogs, docs, tutorials to grow the LLM eval community. 🎯 Requirements Open Source champion and community-driven development. Creative problem solver for ambiguous challenges. Data and metrics driven with empirical results. Technically curious about LLM architectures. Builder mindset: prototyping to production. Hands-on LLM experience and strong TypeScript proficiency. 🎁 Benefits Shape the future of AI evaluation. High impact, real ownership. Fully remote, flexible environment with optional in-person offices. WFH monthly coworking stipend. Work on large-scale production AI workloads.