Mercor is a platform connecting elite creative and technical talent with leading AI research labs. We are currently seeking an experienced AI Safety Specialist to join our team in a fully remote capacity to help ensure the development of safe and reliable artificial intelligence systems.
Key responsibilities
- Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
- Review content involving sensitive domains such as misinformation, political persuasion, self-harm, violence, cyber, and biosecurity.
- Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
- Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
- Provide structured feedback to improve model alignment and safety performance while collaborating with AI researchers.
Requirements
- Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
- Minimum of 5 years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, or security.
- Excellent written English skills combined with strong critical thinking and analytical reasoning.
- Proven ability to consistently evaluate nuanced and policy-sensitive scenarios.
- Familiarity with safety policies, content moderation, or evaluation rubric development is highly preferred.
What we offer
- The opportunity to work on cutting-edge AI safety initiatives with top-tier research labs.
- A fully remote work environment that supports professional growth and collaboration.
- Engagement in high-impact projects that shape the future of responsible AI development.