AI Safety Expert – English, Norwegian

Posted 1ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

AI safety red team expert probing Mercor’s partner AI models for vulnerabilities. Generating adversarial datasets, attack cases, and reproducible safety reports.

Responsibilities:

  • Red team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • Annotate failures, classify vulnerabilities, and flag systemic risks
  • Follow taxonomies, benchmarks, and playbooks to keep testing consistent
  • Produce reproducible reports, datasets, and attack cases
  • Identify vulnerabilities automated tests miss
  • Expand evaluation coverage and reduce production surprises
  • Strengthen customer AI systems through human data-driven red teaming

Requirements:

  • Fluent/native fluency in English and Norwegian
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
  • Ability to probe AI systems adversarially and push systems to breaking points
  • Experience using frameworks or benchmarks for structured testing
  • Ability to explain risks clearly to technical and non-technical stakeholders
  • Adaptability across projects and customers
  • Independent contractor status
  • Must not be an H1-B or STEM OPT candidate
  • Nice-to-have specialties: adversarial ML, cybersecurity, socio-technical risk, or creative probing

Benefits:

  • Fully remote role
  • Flexible schedule; work can be completed on your own schedule
  • Weekly payments via Stripe or Wise
  • Optional participation in higher-sensitivity projects
  • Clear content guidelines and wellness resources
  • Opportunity to build experience in human data-driven AI red teaming
  • Direct role in making AI systems more robust, safe, and trustworthy
  • Collaboration with leading researchers
  • Referral payments of up to $250 per successful referral