Remotery

Senior Site Reliability Engineer

Posted Jun 16

This is a fully remote position, open to applicants in United States.

📋 Description

• Design, scale, and manage robust, cloud-native infrastructure within AWS, focusing on EKS, IAM, RBAC, and contemporary security-first methodologies.

• Develop and enhance CI/CD pipelines utilizing GitHub Actions and GitHub Advanced Security, facilitating speed without sacrificing safety.

• Take ownership of observability throughout the stack using Datadog (metrics, logging, alerting, and tracing).

• Create and maintain Terragrunt, Terraform modules, and infrastructure-as-code (IaC) automation.

• Build internal tools and scripts using Python to automate operational processes and lessen manual tasks.

• Document everything from runbooks to standards to ensure team alignment and system stability.

• Actively engage in Agile workflows with Jira, ensuring clear tracking of tasks, priorities, and progress.

• Participate in on-call rotations, postmortems, and continuous improvement initiatives—maintaining a blameless, team-oriented approach.


⛳️ Requirements

• Over 4 years of experience in a Senior SRE or DevOps position supporting large-scale production cloud infrastructure, ideally within SaaS, PaaS, high-growth, or dynamic environments.

• Extensive experience with AWS (IAM, EKS, VPC, EC2, Secrets Manager, Serverless) and RBAC.

• Familiarity with compliance standards such as HIPAA, HITRUST, or SOC 2.

• Proficient with Terraform, Terragrunt, Helm, and container orchestration technologies.

• Demonstrated experience in developing and managing GitHub Actions for CI/CD, including GitHub Advanced Security features such as secret scanning and code policy enforcement.

• Strong experience with Datadog, including building dashboards, fine-tuning alerts, configuring monitors, and analyzing telemetry data.

• Solid experience in Python scripting for automation and internal tool development.

• You appreciate thorough, accurate documentation as an essential aspect of engineering rather than an afterthought.

• Comfortable working in Agile/Scrum settings with well-managed Jira workflows.

• Practical experience in resource analysis and infrastructure optimization.


🏝️ Benefits

• Eligible for Annual Bonus

• Healthcare benefits, short/long-term disability coverage, life insurance, and 401k

• Paid Parental Leave

• Nine paid holidays & Unlimited PTO

• Remote working arrangements

People also viewed

The CodestJul 26

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
IRIUMJul 26

Ingeniero/a Cloud DevOps

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€33k – €40k/year
ApplyView job
SólidesJul 26

Senior DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ResilincJul 25

Junior/Senior Site Reliability Engineer – Night Shift

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity GroupJul 25

Senior SRE / DevOps Engineer

Anywhere in the WorldFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
HOESSLER & HOESSLERJul 25

DevOps Software Engineer – Career Ambitions

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€65k – €75k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers