AI Safety Specialist Fully Remote
Job Summary
About the job
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco our investors include Benchmark General Catalyst Peter Thiel Adam DAngelo Larry Summers and Jack Dorsey.
Position: AI Safety Red Teamer
Type: Contract
Compensation: $70$84/hour
Location: Remote
Role Responsibilities
- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks unsafe behaviors hallucinations and policy failures.
- Evaluate model robustness across misinformation cyber biosecurity fraud political content and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with AI researchers to improve model alignment robustness and safety.
Qualifications
Must-Have
- Bachelors degree or higher in Computer Science Cybersecurity Journalism Communications Psychology Biology Chemistry Public Policy or a related discipline.
- 5 years of professional experience in AI Safety AI Red Teaming Trust & Safety cybersecurity investigative journalism life sciences or a related field.
- Strong analytical reasoning prompt design and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems.
Preferred
- Experience with AI Red Teaming RLHF SFT AI Alignment or Trust & Safety.
- Familiarity with jailbreak testing prompt engineering or adversarial evaluation methodologies.
- Expertise in one or more grey-area domains including cyber biosecurity political content misinformation or scientific safety.
Application Process (Takes 2030 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form
Resources & Support
- For details about the interview process and platform information please check:
- For any help or support reach out to:
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.