AI Safety Expert, English, Finnish
Posted 1ds ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
AI safety red team expert probing Mercor’s frontier AI models for vulnerabilities. Generating adversarial datasets, reports, and attack cases to improve customer AI safety.
Responsibilities:
- Red team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
- Annotate failures, classify vulnerabilities, and flag systemic risks
- Follow taxonomies, benchmarks, and playbooks to keep testing consistent
- Produce reproducible reports, datasets, and attack cases for customers
- Uncover vulnerabilities that automated tests miss
- Expand evaluation coverage across more scenarios
- Help Mercor customers strengthen the safety and robustness of their AI systems
Requirements:
- Fluent language skills required: English and Finnish
- Native fluency in English and Finnish is required
- Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
- Ability to probe AI systems adversarially and push systems to breaking points
- Experience using frameworks or benchmarks for structured testing
- Ability to explain risks clearly to technical and non-technical stakeholders
- Ability to adapt across projects and customers
- Nice-to-have specialties: adversarial ML, cybersecurity, socio-technical risk, or creative probing
- Independent contractor engagement
- H1-B and STEM OPT candidates are not supported
Benefits:
- Fully remote role that can be completed on your own schedule
- Weekly payments via Stripe or Wise based on services rendered
- Projects can be extended, shortened, or concluded early depending on needs and performance
- Participation in higher-sensitivity projects is optional
- Clear guidelines and wellness resources for higher-sensitivity projects
- Competitive pay
- Referral bonuses of up to $250 per successful referral








