Senior Production Engineer

Posted 14ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Senior Production Engineer leading scalable cloud infrastructure for WP Engine’s WordPress platform. Managing reliability, automation, incidents, and engineering development at internet scale.

Responsibilities:

  • Lead a team of engineers using a Site Reliability Engineering (SRE) mentality
  • Craft, build, and maintain complex systems at scale
  • Curate an engineering culture biased toward shipping while maintaining high standards of stability, performance, security, and scalability
  • Monitor, alert, investigate, and resolve production infrastructure challenges across public cloud environments, including GCP and AWS
  • Identify opportunities and implement automated, scalable solutions to optimize code and operational tasks and reduce toil
  • Serve as a technical point of contact during critical issues
  • Facilitate cross-functional communication, track alert trends, and perform root cause analysis (RCA) to drive structural improvements
  • Execute production changes and security requests, including code deployments and patching pipelines
  • Partner with Product Management and Engineering Management to align technical solutions and execution strategies with business goals and roadmap priorities
  • Coach, mentor, and develop engineers across the organization

Requirements:

  • 5+ years of software engineering, DevOps, or Site Reliability Engineering (SRE) experience building and running production-ready configurations at internet scale
  • Strong history of system architecture, distributed microservices, and self-healing systems with efficient resource and network utilization
  • Advanced expertise executing containerized workloads using Kubernetes orchestration and Docker across public cloud hosting providers
  • Proficient in structural and automated scripting using object-oriented languages
  • Strong experience in Python and Go preferred
  • In-depth understanding of the modern web serving stack and infrastructure automation tools, including Linux, Nginx/Apache, MySQL, PHP, Ansible, and Terraform
  • Hands-on familiarity implementing metrics, monitoring, and alerting systems such as Prometheus or EFK stacks
  • Superb analytical, troubleshooting, and root cause analysis capabilities
  • Ability to work in a team-first environment

Benefits:

  • Company Stock Options (Every employee is an owner in the company)
  • Superannuation Program
  • Employee Assistance Program
  • Supplemental Maternity & Paternity Pay
  • Generous Vacation Time (Who doesn't like time off)
  • One-time $745 AUS Home Office Stipend
  • Company Wellness Days