AI Safety Specialist - Fully Remote | Upto $84/hr

Overview of the Position

The position advertised is for an AI Safety Specialist at Mercor, a company that connects top talent with leading AI research labs. This role is fully remote, and the compensation ranges from $70 to $84 per hour, making it an attractive opportunity for candidates with the right skill set. Mercor is headquartered in San Francisco, and it has backing from prominent investors such as Benchmark, General Catalyst, and others.

Role Responsibilities

As an AI Safety Specialist, your main responsibilities will include:

  • Designing adversarial prompts to test and stress various AI models.
  • Identifying vulnerabilities such as jailbreaks, unsafe behaviors, and hallucinations in AI systems.
  • Evaluating the robustness of models against various types of issues including misinformation, cybersecurity threats, and fraud.
  • Documenting identified vulnerabilities and contributing to reports on safety and red-teaming activities.
  • Collaborating with AI researchers to improve fidelity and safety in model development.

This position requires a proactive approach to identify and mitigate risk factors in AI, making it suitable for individuals who are engaged in cutting-edge technology and safety evaluations.

Required Skills

To be considered for this role, candidates must meet the following qualifications:

  • A Bachelor's degree or higher in fields such as Computer Science, Cybersecurity, Journalism, Communications, Psychology, or any related discipline.
  • At least 5 years of professional experience in relevant fields including AI Safety, Red Teaming, Trust & Safety, or cybersecurity. This can also include experience in investigative journalism or life sciences.
  • Strong skills in analytical reasoning, prompt design, and written communication to convey complex findings effectively.
  • Specific experience in designing adversarial prompts or evaluating advanced AI systems is essential.

Preferred qualifications include:

  • Hands-on familiarity with AI Red Teaming practices, Reinforcement Learning from Human Feedback (RLHF), and Safe Fine-Tuning (SFT).
  • Expertise in various grey-area domains such as misinformation, biosecurity, or political content. This includes experience with jailbreak testing and adversarial evaluation methodologies.

Application Process

The application process is designed for efficiency and takes approximately 20-30 minutes to complete. Interested candidates are required to:

  • Upload their resume.
  • Complete an AI interview that is tailored based on the review of their resume.
  • Fill out a submission form to finalize their application.

It's crucial to approach this application with thorough detail and clarity, as the hiring team reviews applications daily.

Resources and Support

For those who need help during the application process, Mercor provides support channels to assist candidates. It’s recommended to look into all resources available about the interview process and prepare accordingly to enhance the chances of success in securing the role.

Conclusion

This role is an excellent opportunity for professionals looking to delve into the field of AI safety and work on pressing issues related to AI ethics and security. With a strong emphasis on collaboration and innovation at Mercor, the AI Safety Specialist position is set within a vibrant and forward-thinking environment. Candidates with the right mix of skills in technology, communication, and safety protocols should consider applying to be part of this essential aspect of AI development.



This job offer was originally published on himalayas.app

mercor

Estonia

SEO

Contract

August 13, 2026

0 views

0 clicks on Apply Now


Similar job offers


This job offer summary has been generated using automated technology. While we strive for accuracy, it may not always fully capture the nuances and details of the original job posting. We recommend reviewing the complete job listing before making any decisions or applications.