Data Type
Task Types
Subject Matter / Industry
Language
Share link
In this fully remote, hourly contractor role, you will review AI-generated model responses and generate safety-focused evaluation content in English and French. You will be responsible for scoring, annotating, and evaluating AI outputs, focusing on clear reasoning, policy alignment, and safety. Tasks include curating red-team training cases across nuanced, potentially explicit content areas to help prevent toxic or unsafe outputs. Your expertise will help establish robust labeling and safety standards to improve leading AI models.