Job Description
📋 Description Run Seb evaluation and quality operations; manage eval sets and queues. Own the feedback loop between real-world agent performance and the roadmap. Build and refine the ops infra for Seb quality at scale. Own adoption KPIs; understand why people use Seb. Spot patterns across evals and staff feedback indicating product/trust issues. Keep AI leadership informed on program status, quality trends, and risk. 🎯 Requirements 5+ years in product operations or program management, with AI/ML exposure. Comfortable with evals, annotation tools, and quality frameworks; ML background not required. Highly organized; can run multiple workstreams without losing the thread. Clear communicator who can translate technical detail into simple updates. Genuinely curious about AI and how agents behave in production. Comfortable in ambiguity; you build process where none exists yet. 🎁 Benefits Flexible PTO, top-tier health/dental/vision, a 401(k) plan, and meaningful equity. Work location: San Francisco — in-office most days at our FiDi SF office.