Machine Learning Engineer – Model Evaluation, Experimentation

Posted 16hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Machine learning engineer authoring experiments and evaluating frontier AI models for a leading AI lab. Implementing ML and reinforcement-learning tasks, analyzing results, and identifying model limitations remotely in the United States.

Responsibilities:

  • Design well-defined, multi-step machine-learning tasks from real ML research ideas
  • Implement changes, run training experiments, and analyze results to establish correct solutions
  • Build tasks around reinforcement-learning concepts such as reward functions and training behavior
  • Evaluate frontier models and identify where and why they fall short
  • Compare notes with researchers and fellow experts to maintain consistency, rigor, and fairness
  • Work in a tight feedback loop with the lab’s researchers

Requirements:

  • MSc or PhD in machine learning, computer science, or another STEM field, or equivalent practical experience in a research-heavy domain
  • 1+ years of experience in a research or research-engineering role
  • Hands-on experience training and evaluating ML models and running experiments end-to-end, including setup, execution, and analysis
  • Strong familiarity with large language models, including their capabilities, limitations, and evaluation techniques
  • Working proficiency in Python and Git
  • Comfort working in scripting and notebook environments
  • Basic understanding of reinforcement learning, including reward functions and policy training, preferred
  • Past experience in AI training, model evaluation, or benchmark/task authoring preferred
  • High attention to detail and creativity in task design
  • Strong written communication skills
  • Ability to work independently through ambiguous, open-ended problems
  • Ability to engage reliably for approximately 35 hours per week

Benefits:

  • W-2 employment
  • Payroll, benefits, and compliance administered by Cincinnatus LLC
  • Fully remote work within the United States
  • Approximately 35 hours per week
  • Opportunity to be placed at a leading AI lab as part of its extended workforce
  • Reasonable accommodations for qualified individuals with disabilities and disabled veterans throughout the job application process