
Site Reliability Engineer III, DevEx
Posted Jul 30

Posted Jul 30
This is a fully remote position, open to applicants in Massachusetts.
β’ Design and enhance developer experience tools and CI/CD pipelines to guarantee that cloud software updates are deployed reliably and efficiently.
β’ Take ownership of and oversee developer self-service tools and operational processes related to secrets management, deployment rollbacks, and best practices for infrastructure.
β’ Collaborate with cloud engineering teams to diagnose complex infrastructure problems, monitor production health, and minimize operational friction.
β’ Engage in the Cloud SRE weekly on-call rotation to uphold production availability and address system incidents under pressure.
β’ Develop efficient, production-ready automation code and tests in Go and TypeScript to facilitate cloud developer experience workflows.
β’ Proven experience in writing production-quality code in Go or TypeScript along with automated testing frameworks.
β’ Practical expertise with cloud-native infrastructure tools such as Kubernetes, Helm, Terraform, GitHub Actions, and AWS infrastructure.
β’ Solid understanding of observability, distributed troubleshooting, CI/CD improvement, and disaster recovery principles.
β’ Ability to manage context switching effectively and work asynchronously in a high-trust, collaborative environment.
β’ Flock Stock Options
DATAGROUP
Ambush
DuoKey
TEKsystems
Get handpicked remote jobs straight to your inbox weekly.