Senior Engineering Manager – Scalable Machine Learning

Posted 14hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Engineering manager leading scalable machine-learning infrastructure for Latitude AI’s Ford autonomy roadmap. Setting cross-team strategy for data generation, orchestration, distributed training, and MLOps.

Responsibilities:

  • Lead and mentor a team of software engineers developing scalable machine learning infrastructure.
  • Define and own the multi-quarter technical strategy for data generation, orchestration, and distributed training.
  • Represent the Scalable ML team in organization-level planning and architecture decisions.
  • Resolve cross-team tradeoffs involving GPU scheduling and quota, job orchestration, storage, and compute cost.
  • Drive advanced MLOps practices, including model artifact management, experiment tracking, and automated model deployment and testing.
  • Build technical partnerships with ML Platform, Jobs Platform, Cloud Platform, and HPC teams.
  • Maintain high reliability standards and ensure cross-team incidents are root-caused, fixed, and prevented.
  • Manage the Scalable ML team and serve as a technical bridge across organizations.

Requirements:

  • A Bachelor's and 10 years experience or Master's Degree and 8 years experience in Computer Engineering, Computer Science, Electrical Engineering, Robotics, or a related field.
  • Minimum of 10 years of experience in software development, with at least 5 years in a Staff/Principal engineering or senior technical leadership role.
  • A track record of setting technical strategy and direction across multiple teams or organizations.
  • Extensive experience designing, building, and operating large-scale distributed systems and cloud infrastructure for machine learning, including shared compute dependencies such as GPU scheduling/quota, job orchestration, and multi-tenant clusters.
  • Proven expertise in developing and optimizing machine learning data pipelines, including data generation, management, and distributed training.
  • Strong proficiency in Python for developing high-performance, production-quality software.
  • Demonstrated ability to influence and align peer teams and leaders without direct authority.
  • Track record of delivering complex, multi-org software initiatives from conception to production, emphasizing scalability, reliability, and organizational alignment.
  • Candidates must be legally authorized to work in the United States on a permanent basis.
  • Visa sponsorship is available for this position.

Benefits:

  • Competitive compensation packages
  • High-quality individual and family medical, dental, and vision insurance
  • Health savings account with available employer match
  • Employer-matched 401(k) retirement plan with immediate vesting
  • Employer-paid group term life insurance and the option to elect voluntary life insurance
  • Paid parental leave and Adoption/Surrogacy support program
  • Paid medical leave
  • Unlimited vacation and 15 paid holidays
  • Daily lunches, snacks, and beverages available in all office locations
  • Pre-tax spending accounts for healthcare and dependent care expenses
  • Pre-tax commuter benefits
  • Monthly wellness stipend
  • Backup child and elder care program
  • Professional development reimbursement
  • Employee assistance program
  • Discounted programs that include legal services, identity theft protection, pet insurance, and more
  • Company and team bonding outlets: employee resource groups, quarterly team activity stipend, and wellness initiatives
  • Annual bonus programs
  • Equity compensation