AI Safety Specialist - Fully Remote Contract Position
Remote · Part Time
Be the first to apply
- Experience
- 5+ yrs
- Salary
- USD 70 – USD 84 / hour
- Openings
- 1
- Posted
- 1 hour ago
- Work mode
- Work from home
- Education
- Bachelor's degree or higher
- Resume
- Required to apply
Job description
About the Role
Mercor connects top-tier creative and technical experts with leading AI research laboratories. Based in San Francisco, the company is backed by major investors such as Benchmark, General Catalyst, and prominent figures including Peter Thiel and Jack Dorsey. This opportunity is for a contract AI Safety Red Teamer working remotely, with compensation ranging from $70 to $84 per hour.
Core Responsibilities
- Create adversarial prompts aimed at stress-testing cutting-edge AI models.
- Detect vulnerabilities such as jailbreaks, unsafe actions, hallucinated information, and failures in policy adherence.
- Assess model resilience concerning misinformation, cybersecurity threats, biosecurity, fraud, political content, and other critical sensitive areas.
- Document discovered weaknesses and assist in generating safety benchmarking and red team reports.
- Collaborate closely with AI researchers to enhance model alignment, robustness, and overall safety.
Qualification Criteria
- Minimum bachelor's degree in disciplines such as Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or related fields.
- At least 5 years of professional experience in areas like AI Safety, Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or relevant domains.
- Strong analytical thinking, expertise in prompt design, and excellent written communication abilities.
- Hands-on experience with adversarial prompt design or evaluating advanced AI systems.
Preferred Background
- Experience in AI Red Teaming, Reinforcement Learning from Human Feedback (RLHF), Supervised Fine Tuning (SFT), AI Alignment, or Trust & Safety roles.
- Knowledge of jailbreak testing, prompt engineering, or adversarial evaluation techniques.
- Specialization or deep familiarity with grey-area sectors such as cybersecurity, biosecurity, political content, misinformation, or scientific safety.
Application Process
- Submit your resume.
- Complete an AI-driven interview based on your resume.
- Fill out the application form.
Applications are reviewed daily. To be considered, ensure all application steps including the AI interview are completed promptly.
Additional Support
- For interview details and platform guidance, contact the employer directly.
- Assistance and support are available via the provided contact email.
Skills
Work styles they’re looking for
Written Communication