Remotely
ai evaluationdata annotationcontent moderationdatapolicy operationsclaudechatgptgeminirlhf
Job Description
📋 Description
- Learn new customer policies, definitions, taxonomies, and evaluation rubrics quickly
- Evaluate user requests and AI model responses within the full relevant conversation context
- Distinguish between closely related labels, severity levels, and policy boundaries
- Select the most defensible classification when a case is genuinely ambiguous
- Write concise, evidence-based rationales that cite relevant policy language and conversation details
- Identify policy gaps, contradictions, unclear definitions, and emerging edge cases
🎯 Requirements
- Enjoy making precise distinctions between cases that other people might consider equivalent
- Notice when one word, contextual detail, or change in intent materially affects the answer
- Hold a strong opinion without becoming attached to being right
- Explain judgment calls clearly enough that another person can audit your reasoning
- Ask productive questions when a policy is ambiguous instead of guessing or forcing certainty
- Separate personal views from the standard a customer has asked you to apply
🎁 Benefits
- Shape how every career evolves in the AI economy, at global scale, with impact your friends, family
- Partner hand-in-hand with world-class AI labs, Fortune 500 partners and the world's top educational
- Work together with engineers, scientists, operators, and more from Palantir, Meta, Scale AI, and
- Build a massive, fast-growing business with billions in revenue
Back to all jobs