Senior Software Engineer, DevOps

Posted 11hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Senior DevOps Engineer building cloud infrastructure, CI/CD, and observability for Atria’s preventive healthcare platform. Improving reliability and developer productivity across engineering teams.

Responsibilities:

  • Design, build, and maintain Google Cloud Platform infrastructure using Terraform.
  • Define infrastructure patterns and standards.
  • Own and improve GitHub Actions CI/CD pipelines.
  • Identify and eliminate operational toil through scalable automation and tooling.
  • Establish monitoring, dashboards, and alerting in Datadog, and reduce alert noise.
  • Lead incident response within the on-call rotation and drive postmortems and follow-ups.
  • Define and track Service Level Objectives for core infrastructure and build systems.
  • Partner with product engineering teams on deployment pipelines, environment issues, and build troubleshooting.
  • Own preview and staging environments, including data sync, masking, and cleanup routines.
  • Improve developer experience through tooling, runbooks, and documentation.
  • Write tested and reviewed infrastructure code.
  • Lead design reviews and RFCs with an operations-focused perspective.
  • Design systems balancing reliability, performance, and security.
  • Conduct code reviews and mentor engineers through pairing, knowledge sharing, and documentation.
  • Collaborate with the Tech Lead, DevOps, Platform, Data Engineering, Clinical Experience, Member Experience, and Care Delivery teams.

Requirements:

  • ~5+ years of professional experience in DevOps, SRE, infrastructure, or backend engineering in production environments.
  • Hands-on experience designing and operating infrastructure in at least one cloud provider, ideally Google Cloud Platform.
  • Track record of owning and shipping automation, pipelines, or infrastructure that made a team measurably more productive or reliable.
  • An enthusiasm for developer productivity and making teams as impactful as possible.
  • Deep experience with infrastructure-as-code, ideally Terraform, and building CI/CD pipelines, ideally GitHub Actions.
  • Proficient software engineering ability and strong command of Linux.
  • Strong instincts for monitoring and observability, with confident debugging across logs, traces, and metrics.
  • Solid experience with relational databases, including MySQL and PostgreSQL, and containerized workloads.
  • Strong grounding in reliability, performance, and security fundamentals, with judgment to make sound tradeoffs.
  • Nice to have: experience in healthcare, digital health, or regulated domains such as HIPAA, PHI, and SOC 2.
  • Nice to have: experience with containers and orchestration, including Docker and Kubernetes.
  • Nice to have: exposure to leading incident response, on-call, and postmortem practices.
  • Nice to have: experience with database migrations or managing multiple environments at scale.