Senior Site Reliability Engineer, AWS Cloud

Posted Aug 10

This is a fully remote position, open to applicants in Romania.

📋 Description

• Design and lead the development of reliability, scalability, and performance for the multi-cloud provisioning platform across all production environments.

• Create and implement comprehensive automation pipelines to remove manual processes and minimize technical burdens.

• Establish, track, and enhance essential system health metrics (SLIs/SLOs), such as latency, throughput, error rates, and capacity utilization.

• Provide data-informed architectural suggestions.

• Facilitate collaboration with product and cross-functional engineering teams to incorporate reliability and security aspects early in the Software Development Life Cycle (SDLC).

• Develop effective incident detection systems and oversee incident management.

• Address critical incidents and conduct in-depth Root Cause Analysis (RCA).

• Implement long-term engineering solutions to prevent future issues.

• Review platform operations to pinpoint bottlenecks, eliminate single points of failure, and streamline structural complexity.


⛳️ Requirements

• Over 8 years of experience in cloud environments, specifically in SRE or Cloud Platform/Reliability Engineer roles.

• Significant experience in cloud development and multi-cloud settings.

• Strong familiarity with AWS cloud.

• Understanding of cloud architecture, scalability, and high-availability design principles.

• Practical experience with Kubernetes and container orchestration.

• Proficiency in Terraform and Infrastructure as Code (IaC).

• Experience in designing automation and CI/CD pipelines to alleviate operational burdens.

• Deep understanding of SRE principles, SLIs/SLOs, monitoring, and observability.

• Demonstrated experience with Incident Management, Root Cause Analysis (RCA), and reliability engineering.

• Capability to identify performance bottlenecks, single points of failure, and architectural vulnerabilities.


🏝️ Benefits

• Competitive salary and performance-based bonuses.

• Comprehensive health, dental, and vision insurance packages.

• Flexible work hours and remote work options.

• Opportunities for professional development and continuous learning.

• Supportive and inclusive work culture.

People also viewed

FourEnergy GmbH17 hours ago

Senior DevOps Engineer – Operations

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ICF19 hours ago

Lead DevOps Engineer

US flagVirginia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$131.3k – $223.1k/year
ApplyView job
Mastercam23 hours ago

DevSecOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
C&S Informática1 day ago

DevOps Engineer – Freelance/Contract, Mid-Level/Senior

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Convene1 day ago

Support and Deployment Engineer

SA flagSaudi Arabia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity Group1 day ago

SRE Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers