AI Safety Experts – English, Norwegian
Posted 1ds ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
AI safety red team expert probing Mercor’s partner AI models for vulnerabilities. Generating adversarial datasets, attack cases, and reproducible safety reports.
Responsibilities:
- Red team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
- Annotate failures, classify vulnerabilities, and flag systemic risks
- Follow taxonomies, benchmarks, and playbooks to keep testing consistent
- Produce reproducible reports, datasets, and attack cases
- Probe AI systems for vulnerabilities that automated tests miss
- Expand evaluation coverage and reduce production surprises
- Strengthen customer AI systems through human data-driven red teaming
Requirements:
- Native fluency in English and Norwegian
- Prior red teaming experience, including AI adversarial work, cybersecurity, or socio-technical probing
- Ability to probe AI systems adversarially and push systems to breaking points
- Experience using frameworks or benchmarks for structured testing
- Ability to explain risks clearly to technical and non-technical stakeholders
- Ability to adapt across projects and customers
- Independent contractor status
- H1-B and STEM OPT candidates are not supported
Benefits:
- Fully remote work
- Flexible schedule / work on your own schedule
- Weekly payments via Stripe or Wise
- Projects may be extended, shortened, or concluded early depending on needs and performance
- Wellness resources for higher-sensitivity projects
- Reasonable accommodations upon request
- Up to $250 referral payment for each successful referral








