
Site Reliability Engineer, Cloud Infrastructure
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in Utah.
• Develop, sustain, and enhance the cloud infrastructure that supports Weave's services.
• Guarantee the reliability, scalability, and performance of the platform.
• Automate routine infrastructure tasks.
• Design and establish highly available and scalable systems.
• Ensure the seamless day-to-day functioning of Weave's infrastructure.
• Create and refine tools and standards for automation, scaling, monitoring, and alerting.
• Work in collaboration with product teams to address production challenges, enhance monitoring, and utilize cloud services.
• Take part in a weekly on-call rotation.
• Report directly to the Engineering Manager.
• Expertise in at least one cloud platform is essential.
• Strong understanding of Kubernetes and Docker.
• Experience with automation tools such as Puppet, Salt, Ansible, and Terraform.
• Proficiency in writing automation scripts using Go, Python, etc.
• Background in designing highly available and scalable systems.
• Proficient with version control systems like Git and CI/CD principles.
• Excellent problem-solving capabilities and a systematic approach to troubleshooting complex issues.
• Enthusiasm for Infrastructure as Code.
• Extensive expertise in Kubernetes, including cluster design, deployment, and ongoing management for large-scale applications.
• Familiarity with advanced GCP services and architectures.
• Experience managing infrastructure and applications using IaC, GitOps, and ArgoCD.
• Successful completion of a background check.
• Legally authorized to work in the United States and capable of providing valid Form I-9 documentation.
• Opportunity for remote work.
• Participation in a weekly on-call rotation.
• Commitment to an equal opportunity and inclusive workplace.
• Provision for disability or special-needs accommodation.
Slate Auto
Funding Xchange
Leidos
LeoLabs
Get handpicked remote jobs straight to your inbox weekly.