Role directory

Policy & Safety jobs

4 active, referral-verified opportunities.

Policy & Safety / Remote

Privacy / regulatory compliance Evaluator

About the role — We are hiring expert Evaluators in Privacy / regulatory compliance to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Privacy / regulatory compliance. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Policy & Safety / Remote

Legal / compliance Evaluator

About the role — We are hiring expert Evaluators in Legal / compliance to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Legal / compliance. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Policy & Safety / Remote

AI Safety Red Teamer

We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area") topics. — Responsibilities • Design adversarial prompts to stress-test frontier AI models. • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures. • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains. • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports. • Collaborate with AI researchers to improve model alignment, robustness, and safety. — Required Qualifications • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline. • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field. • Strong analytical reasoning, prompt design, and written communication skills. • Experience designing adversarial prompts or evaluating frontier AI systems. — Preferred Qualifications • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety. • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies. • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety. — Why Join? • Help secure and strengthen the next generation of frontier AI models. • Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams. • Influence how AI systems respond to complex, real-world safety challenges.

$70 - $84 / hourOpen / Referral verified
Policy & Safety / Remote

AI Safety Practitioner

We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback. — Responsibilities • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality. • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains. • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking. • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations. • Provide structured feedback to improve model alignment and safety performance. • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives. — Required Qualifications • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline. • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field. • Excellent written English, critical thinking, and analytical reasoning skills. • Ability to consistently evaluate nuanced and policy-sensitive scenarios. — Preferred Qualifications • Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation. • Familiarity with safety policies, content moderation, or evaluation rubric development. • Experience reviewing complex, high-risk, or ambiguous content. — Why Join? • Shape the safety and behaviour of frontier AI models used by millions worldwide. • Work on challenging, real-world safety evaluations across nuanced and high-impact domains. • Collaborate with leading AI researchers, engineers, and safety teams.

$60 - $70 / hourOpen / Referral verified