AI Safety Practitioner
About the Role
We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.
What You'll Do
- Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality
- Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains
- Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking
- Identify unsafe outputs, hallucinations, reasoning failures, and policy violations
- Provide structured feedback to improve model alignment and safety performance
- Collaborate with AI researchers and safety teams on ongoing evaluation initiatives
Requirements
- Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline
- 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field
- Excellent written English, critical thinking, and analytical reasoning skills
- Ability to consistently evaluate nuanced and policy-sensitive scenarios
Preferred Qualifications
- Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation
- Familiarity with safety policies, content moderation, or evaluation rubric development
- Experience reviewing complex, high-risk, or ambiguous content
Why Join
- Shape the safety and behaviour of frontier AI models used by millions worldwide
- Work on challenging, real-world safety evaluations across nuanced and high-impact domains
- Collaborate with leading AI researchers, engineers, and safety teams
Compensation & Logistics
- Type: Independent Contractor
- Compensation: Competitive, paid weekly via Stripe or Wise
- Location: Remote
- Application: Submit your resume or relevant professional background. Qualified applicants may be asked to complete a brief assessment or submit additional information. Applications are reviewed on a rolling basis.
Related roles
Writing Expert
Apply your literary expertise — playwriting, fiction, journalism, essays, or social media — to advanced AI training projects with leading AI labs. Flexible, project-based, remote.
Website Designer (UK-Based) — AI Training
Contribute world-class website design expertise to a frontier AI training project — a remote, own-schedule independent contractor role for UK-based designers, up to 40 hours/week.
User/Customer Research and Feedback Synthesis Evaluator
Put your user and customer research expertise to work evaluating AI-generated work products for accuracy, rigor, and quality — a flexible, remote, hourly contract with Mercor.
