Job Overview
We are seeking detail-oriented Trust & Safety Analysts – AI Evaluation to review and evaluate AI-generated outputs using content safety policies, project guidelines, and expert human judgment.
In this role, you will assess complex AI outputs and provide accurate, consistent evaluations that help improve the safety and quality of AI systems. You will review situations where context, language, intent, and cultural understanding are important in determining the appropriate evaluation.
An ideal candidate has strong analytical skills, sound judgment, and experience applying content safety policies consistently. You should be comfortable reviewing complex or ambiguous content and making decisions based on defined guidelines rather than personal opinion.
Your work will provide high-quality human feedback and evaluation data that supports the training and improvement of AI systems.
Project Details
• Contract Type: Freelance, with the potential to convert to a full-time role.
• Pay Rate: US$24 per hour
• Location: United States
• Language: English
Responsibilities
• Review and evaluate AI-generated outputs according to defined content safety policies and project guidelines.
• Apply expert human judgment to assess complex, ambiguous, or context-dependent situations.
• Grade, classify, label, or annotate AI outputs accurately and consistently.
• Consider context, language, intent, and cultural differences when evaluating content.
• Identify potential content safety or policy concerns based on defined guidelines.
• Evaluate difficult cases and make informed decisions when the appropriate outcome depends on multiple factors.
• Provide accurate human feedback to support AI system training and evaluation.
• Identify unclear or unusual cases and flag them according to defined project processes.
• Document decisions and supporting reasoning clearly when required.
• Apply detailed project guidelines consistently across assigned tasks.
• Maintain high levels of quality, accuracy, consistency, and attention to detail.