Senior Site Reliability Engineer
Posted 2ds ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Senior Site Reliability Engineer responsible for infrastructure managing high-volume logistics operations on GCP. Collaborating to enhance reliability, automate processes, and improve monitoring practices.
Responsibilities:
- Own architecture and implementation of scalable, reliable infrastructure on GCP
- Manage containerized workloads on Kubernetes
- Build monitoring, alerting, and observability in Datadog
- Drive automation to improve efficiencies
- Design and maintain CI/CD pipelines in GitHub Actions
- Collaborate with data and development teams
Requirements:
- 5+ years in SRE, platform, or infrastructure engineering
- Strong hands-on experience with GCP core services (GKE, Cloud Run, AlloyDB, networking, IAM)
- Fluent in Docker and Kubernetes
- Deep Terraform experience
- Proficient in a programming language: TypeScript, Python, Go, or similar
- Build monitoring and alerting that's actionable (Datadog or equivalents)
- Understanding of distributed systems fundamentals
- Experience with Git and collaborative development workflows
- Incident management experience
- Strong communication skills
- Ownership and accountability
- Collaborative approach
Benefits:
- Remote work options
- Paid time off

















