
Staff Engineer, Developer Infrastructure
Posted 5 days ago

Posted 5 days ago
This is a fully remote position, open to applicants in United States.
β’ Develop and manage the developer infrastructure for Ava Labs' Platform Engineering teams.
β’ Collaborate with Platform Engineering to gain insights into development, testing, release, observability, and debugging processes.
β’ Enhance CI/CD pipelines utilizing Bazel and GitHub Actions for unit, end-to-end, performance, and release-critical testing.
β’ Optimize the provisioning, inspection, comparison, and decommissioning of both ephemeral and long-running test environments.
β’ Advance observability infrastructure to analyze data across different runs, versions, and configurations.
β’ Strengthen the accuracy and robustness of consensus, networking, and protocol-adjacent systems through comprehensive testing, fuzzing, chaos engineering, and failure-oriented test design.
β’ Transform recurring challenges faced by the team into effective tools, automation, and documentation.
β’ Convey technical context through design reviews, illustrative examples, and technical writing.
β’ Troubleshoot flaky tests, enhance release checks, refine documentation, evaluate tool adoption, and ensure the system remains beneficial.
β’ 7+ years of experience in writing production-level software.
β’ Proficient in Go and/or Rust.
β’ Familiarity with extensive legacy codebases involving multiple teams and shared ownership boundaries.
β’ Experience in building or significantly enhancing developer infrastructure, CI/CD systems, testing frameworks, observability tools, or similar engineering productivity systems.
β’ Strong commitment to testing and an ability to identify and pragmatically alleviate bottlenecks.
β’ Practical knowledge of Linux, containers, and infrastructure-as-code.
β’ Preferred: Experience with GitHub Actions or comparable CI/CD systems; Bazel or similar build systems.
β’ Preferred: Background in distributed systems, blockchain infrastructure, consensus systems, P2P networking, or other high-reliability systems.
β’ Preferred: Familiarity with Kubernetes or similar orchestration frameworks.
β’ Preferred: Experience in large-scale observability including tracing, profiling, and metrics.
β’ Preferred: Utilizing data to inform performance optimization.
β’ Preferred: Knowledge of chaos engineering, fuzzing, property-based testing, and failure-injection systems.
β’ Ability to operate at a Staff level and create or transform workstreams that other teams rely on.
β’ Token and equity package.
β’ Remote work arrangement.
β’ Global team environment.
β’ Equal Opportunity Employer.
Second Nature
Headway
SYNCREON
Rentokil Pest Control North America
Get handpicked remote jobs straight to your inbox weekly.