Principal Operations Engineer, Network

Posted 1ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Principal network operations engineer overseeing hyperscale AI data center networks at Fluidstack. Leading operational readiness, high-risk changes, vendor standards, and fleet-wide incident analysis.

Responsibilities:

  • Serve as the most senior technical authority for the operational network fleet across the hyperscale AI data center portfolio
  • Lead site assessments and operational audits
  • Drive technical readiness of the network operations team ahead of site activation
  • Review network platforms and integration designs from an operational perspective
  • Feed operational learnings back into network engineering, deployment, and supply chain
  • Author, approve, and execute high-risk network MOPs and change records in live production
  • Lead fleet-wide root cause analysis on disruptions through closure
  • Hold OEMs, ODMs, and service vendors accountable to operational standards
  • Act as a connective force across network operations, network engineering, compute operations, facilities, supply chain, and customer-facing teams
  • Travel 50–75%

Requirements:

  • Career operating mission-critical network topologies at scale
  • Significant experience as the senior technical voice on a site, campus, or fleet
  • Experience in IP network operations and optical networking
  • Experience with TCP/IP and network routing protocols including OSPF, IS-IS, BGP, and MPLS
  • Knowledge of physical network infrastructure devices
  • Experience authoring and executing high-risk MOPs
  • Experience leading root cause analysis on significant network events through closure
  • Experience holding OEMs, ODMs, and deployment partners to operational standards
  • Clear technical writing, including health assessments, RCAs, and design feedback
  • Ability to teach and improve the technical readiness of surrounding teams
  • Bonus: hyperscale or large HPC fleets supporting thousands of endpoints
  • Bonus: Linux and hardware management tooling
  • Bonus: standing up new sites from handover to steady state
  • Bonus: scripting for fleet-scale operations
  • Authorized to work in the United States

Benefits:

  • Competitive total compensation package including salary and equity
  • Equity in the form of restricted stock units
  • Retirement or pension plan, in line with local norms
  • Health, dental, and vision insurance
  • Generous PTO policy, in line with local norms