Remotery

Senior Site Reliability Engineer

Posted 14 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Lead intricate infrastructure migrations and projects across various production environments and jurisdictions.

• Develop and sustain platform tooling and automation utilizing ArgoCD, Helm, GitHub Actions, release pipelines, and service onboarding workflows.

• Provide infrastructure consultation, dependency resolution, architecture proposal reviews, and platform-tooling adoption support to development teams.

• Design and enhance observability and alerting through Datadog monitors, dashboards, and runbooks.

• Analyze and troubleshoot production issues through systematic debugging and root cause analysis.

• Engage in on-call rotations to ensure platform reliability.

• Mentor colleagues and play a role in architecture decisions and ongoing enhancement.


⛳️ Requirements

• Over 5 years of experience in a comparable position (DevOps, Site Reliability Engineer).

• Extensive experience in operating and troubleshooting Kubernetes within a production Linux environment, covering cluster lifecycle, networking, storage, and scheduling.

• Familiarity with AWS, GCP, and/or on-premise environments.

• Proficient in at least two programming languages among Go, Python, and Bash/Shell.

• In-depth knowledge of distributed systems, failure modes, networking fundamentals, capacity planning, and performance analysis.

• Experience with GitOps and CI/CD workflows, including ArgoCD, Helm, GitHub Actions, or similar tools.

• Knowledge of infrastructure-as-code practices, including Terraform, Helm, or equivalent technologies.

• Proven history of leading complex migrations or infrastructure projects involving cross-team dependencies.

• Strong skills in incident response and troubleshooting.

• Excellent technical communication and documentation capabilities.


🏝️ Benefits

• Attractive compensation package.

• Comprehensive benefits offering.

• Enjoyable and relaxed work atmosphere.

• Reimbursements for education and conferences.

• Bonus eligibility for most non-sales roles.

• Top-tier benefits for qualified employees.

• Tailored support options, expert guidance, and continuous tools to assist employees physically, financially, and emotionally.

People also viewed

HubSpot12 hours ago

Principal Software Engineer, Developer Acceleration – Release Engineering

IE flagIreland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$242k – $387.2k/year
ApplyView job
InfluxData12 hours ago

DevOps Engineer

GB flagUnited Kingdom, +7 more countriesFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
LocalStack13 hours ago

Senior DevOps Engineer

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€50k – €78k/year
ApplyView job
Amigo Tech14 hours ago

DevOps Engineer, Mid-Level

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Smarthis14 hours ago

Senior Cloud Platform – DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Tandem Diabetes Care14 hours ago

Principal Site Reliability Engineer – Temp to Hire

US flagUnited States OnlyFreelanceDevOps & Site Reliability Engineer (SRE)$165k – $185k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers