Senior Site Reliability Engineer

Posted 2ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Senior Site Reliability Engineer responsible for infrastructure managing high-volume logistics operations on GCP. Collaborating to enhance reliability, automate processes, and improve monitoring practices.

Responsibilities:

  • Own architecture and implementation of scalable, reliable infrastructure on GCP
  • Manage containerized workloads on Kubernetes
  • Build monitoring, alerting, and observability in Datadog
  • Drive automation to improve efficiencies
  • Design and maintain CI/CD pipelines in GitHub Actions
  • Collaborate with data and development teams

Requirements:

  • 5+ years in SRE, platform, or infrastructure engineering
  • Strong hands-on experience with GCP core services (GKE, Cloud Run, AlloyDB, networking, IAM)
  • Fluent in Docker and Kubernetes
  • Deep Terraform experience
  • Proficient in a programming language: TypeScript, Python, Go, or similar
  • Build monitoring and alerting that's actionable (Datadog or equivalents)
  • Understanding of distributed systems fundamentals
  • Experience with Git and collaborative development workflows
  • Incident management experience
  • Strong communication skills
  • Ownership and accountability
  • Collaborative approach

Benefits:

  • Remote work options
  • Paid time off