Staff Engineer, Developer Infrastructure
Posted 2hrs ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Staff Engineer building CI/CD, testing, observability, and developer tooling for Ava Labs’ Avalanche blockchain platform. Improving reliability and engineering velocity across Platform Engineering.
Responsibilities:
- Build and operate developer infrastructure for Ava Labs’ Platform Engineering teams
- Partner across Platform Engineering to understand development, testing, release, observability, and debugging workflows
- Evolve CI/CD pipelines built on Bazel and GitHub Actions across unit, end-to-end, performance, and release-critical tests
- Improve provisioning, inspection, comparison, and retirement of ephemeral and long-running test environments
- Evolve observability infrastructure to compare data across runs, versions, and configurations
- Improve correctness and resilience of consensus, networking, and protocol-adjacent systems through end-to-end testing, fuzzing, chaos engineering, and failure-oriented test design
- Turn recurring team pain points into tools, automation, and documentation
- Share technical context through design reviews, examples, and technical writing
- Debug flaky tests, tighten release checks, improve documentation, review tool adoption, and follow through on system usefulness
Requirements:
- 7+ years writing production software
- Strong ability in Go and/or Rust
- Experience in large legacy codebases with multiple teams and shared ownership boundaries
- Experience building or significantly improving developer infrastructure, CI/CD, test infrastructure, observability, release tooling, or comparable engineering productivity systems
- Strong testing discipline and ability to diagnose and pragmatically relieve bottlenecks
- Working fluency with Linux, containers, and infrastructure-as-code
- Preferred: GitHub Actions or similar CI/CD systems; Bazel or similar build systems
- Preferred: Distributed systems, blockchain infrastructure, consensus systems, P2P networking, or other high-reliability systems
- Preferred: Kubernetes or similar orchestration frameworks
- Preferred: Large-scale observability including tracing, profiling, and metrics
- Preferred: Using data to drive performance optimization
- Preferred: Chaos engineering, fuzzing, property-based testing, and failure-injection systems
- Ability to operate at Staff scope and create or transform workstreams on which other teams depend
Benefits:
- Token and equity package
- Remote work arrangement
- Global team environment
- Equal Opportunity Employer



















