Remote

24-MAG

📍 New York, New York, United States

Full-time computer-and-mathematical

Job Description

We are sharing a specialised part-time consulting opportunity for experienced AI safety, red-teaming, trust and safety, cybersecurity, investigative research, and scientific-risk professionals with expertise in adversarial evaluation of advanced AI systems.

This role supports a frontier AI safety initiative focused on identifying vulnerabilities, unsafe behaviours, policy failures, and robustness gaps through structured adversarial testing. Selected professionals will design challenging prompts, evaluate model behaviour across high-risk and ambiguous scenarios, document findings, and contribute to safety benchmarks used to improve model alignment and reliability.

Key Responsibilities

Adversarial Prompt Design

  • Design sophisticated prompts that stress-test advanced AI systems
  • Develop realistic scenarios intended to expose model limitations, policy weaknesses, and inconsistent behaviour
  • ...
Apply for this Position