AI Safety Specialist
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
AI safety specialist supporting Mercor’s short-term AI safety project. Applying red-teaming and related domain expertise during a fast-paced remote sprint.
Responsibilities:
- Contribute to a short-term AI safety project during the week of Oct 5th, 2026
- Apply domain expertise and strong red-teaming experience to tackle an AI safety challenge
- Work independently and respond reliably during a fast-paced sprint
Requirements:
- Must be based in the U.S.
- Must be available for 3+ hours per day during the week of Oct 5th, 2026
- Reliable availability and responsiveness throughout the fast-paced sprint
- Strong written English and attention to detail
- Experience in AI safety, red teaming, cybersecurity, trust and safety, content moderation, policy evaluation, or adversarial testing
- Comfort working independently under tight deadlines
- Strong red-teaming experience
- This is not an entry-level learning project
Benefits:
- Earn up to $250 for each successful referral
- Reasonable accommodations upon request
Similar Jobs

AI Safety Expert – English, Marathi
AI safety red teamer probing conversational AI models for Mercor, which partners with AI labs and enterprises to train frontier systems. Generating reproducible vulnerability data and attack cases.

AI Safety Expert – English, Tamil
AI safety red teamer probing conversational AI models for Mercor, which partners with leading labs to train frontier systems. Generating vulnerability data and reproducible attack artifacts.

AI Safety Expert, English, Punjabi
AI red-teaming experts probing conversational models for Mercor’s frontier AI safety projects. Generating vulnerability data, attack cases, and reproducible reports for AI labs and enterprises.

AI Safety Red Teamer
AI Safety Red Teamer stress-testing frontier AI models for Mercor. Identifying vulnerabilities and improving model alignment, robustness, and safety through adversarial evaluation.

AI Safety Red Teamer
AI Safety Red Teamer stress-testing frontier AI models for Mercor. Identifying vulnerabilities and improving model alignment, robustness, and safety.

Senior/Staff AI Model Engineer
AI model engineer owning evaluation, safety, and reliability for Clear Street’s AI copilot. Improving model quality across trading workflows on a cloud-native brokerage platform.

Trainer for CAD and Artificial Intelligence in Engineering
CAD- und KI-Trainer:in für eine Qualifizierung zur modernen digitalen Produktentwicklung. Vermittlung von Generative Design, KI-gestützten CAD-Workflows und technischer Bewertung.

Trainer – CAD, Artificial Intelligence & Engineering
Trainer:in für CAD und KI im Engineering bei einem digitalen Weiterbildungsanbieter. Vermittlung praxisnaher KI-gestützter Konstruktions- und CAD-Workflows an Maschinenbau-Fachkräfte.

Senior Director Analyst – AI Technology Economics
Senior Gartner analyst shaping AI economics research across technology markets. Developing thought leadership and advising executives through data-driven economic insights.

AI Safety Expert – English, Kannada
AI Safety Experts red-teaming conversational models for Mercor, which partners with AI labs to train frontier systems. Probing vulnerabilities and producing reproducible safety data.

AI Safety Expert – English, Kannada
AI safety red teamer probing conversational AI models for Mercor, which partners with leading AI labs to train frontier systems. Creating vulnerability data, attack cases, and reproducible reports.

AI Safety Expert – English, Marathi
AI safety red-teamer probing Mercor’s frontier AI models with adversarial English and Marathi inputs. Generating vulnerability annotations, attack cases, datasets, and reproducible reports.

AI Safety Red Teamer
AI Safety Red Teamer designing adversarial prompts and uncovering vulnerabilities in Mercor’s frontier AI projects. Evaluating model behavior across high-risk domains and supporting AI safety research.

Staff AI Builder
Staff AI Builder designing and shipping production GenAI systems for Robots and Pencils’ enterprise clients. Leading agentic AI, RAG, AWS architecture, reliability, and technical mentoring.

AI Benchmark Engineer, Native Language Specialist – Chinese
AI benchmark engineer building Chinese-language Terminal-Bench tasks for Lilt’s multilingual AI platform. Validating coding agents, reference implementations, and deterministic verifiers.

AI Benchmark Engineer, Native Language Specialist – Chinese
AI benchmark engineer creating Chinese-language Terminal-Bench tasks for Lilt’s multilingual AI platform. Validating coding agents, verifier scripts, and multilingual software robustness.

AI Benchmark Engineer, Native Language Specialist – Arabic
Arabic AI benchmark engineer designing Terminal-Bench tasks for Lilt’s multilingual AI and human-verified language services. Building multilingual assets, verifiers, and calibration workflows for coding agents.

AI Benchmark Engineer, Native Language Specialist – Arabic
Arabic-speaking benchmark engineer designing Terminal-Bench tasks for LILT’s multilingual AI evaluation suite. Building datasets, verifiers, and calibrated tests for coding agents.

AI Safety Experts – English, Telugu
AI safety red-teamer probing conversational models for Mercor’s human-data AI training projects. Generating vulnerability datasets, attack cases, and reproducible reports to strengthen frontier AI systems.

AI Safety Expert – English, Tamil
AI red-team expert probing conversational models for Mercor, which partners with AI labs to train frontier systems. Generating adversarial data, vulnerability reports, and reproducible attack cases.
