Remotery

Senior Site Reliability Engineer, AWS Cloud

Posted 3 hours ago

This is a fully remote position, open to applicants in Romania.

📋 Description

• Design and lead the development of reliability, scalability, and performance for the multi-cloud provisioning platform across all production environments.

• Create and implement comprehensive automation pipelines to remove manual processes and minimize technical burdens.

• Establish, track, and enhance essential system health metrics (SLIs/SLOs), such as latency, throughput, error rates, and capacity utilization.

• Provide data-informed architectural suggestions.

• Facilitate collaboration with product and cross-functional engineering teams to incorporate reliability and security aspects early in the Software Development Life Cycle (SDLC).

• Develop effective incident detection systems and oversee incident management.

• Address critical incidents and conduct in-depth Root Cause Analysis (RCA).

• Implement long-term engineering solutions to prevent future issues.

• Review platform operations to pinpoint bottlenecks, eliminate single points of failure, and streamline structural complexity.


⛳️ Requirements

• Over 8 years of experience in cloud environments, specifically in SRE or Cloud Platform/Reliability Engineer roles.

• Significant experience in cloud development and multi-cloud settings.

• Strong familiarity with AWS cloud.

• Understanding of cloud architecture, scalability, and high-availability design principles.

• Practical experience with Kubernetes and container orchestration.

• Proficiency in Terraform and Infrastructure as Code (IaC).

• Experience in designing automation and CI/CD pipelines to alleviate operational burdens.

• Deep understanding of SRE principles, SLIs/SLOs, monitoring, and observability.

• Demonstrated experience with Incident Management, Root Cause Analysis (RCA), and reliability engineering.

• Capability to identify performance bottlenecks, single points of failure, and architectural vulnerabilities.


🏝️ Benefits

• Competitive salary and performance-based bonuses.

• Comprehensive health, dental, and vision insurance packages.

• Flexible work hours and remote work options.

• Opportunities for professional development and continuous learning.

• Supportive and inclusive work culture.

People also viewed

Endava3 hours ago

Senior DevOps Engineer, Dynatrace

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Jones Lang LaSalle Americas, Inc.3 hours ago

Reliability Engineer

US flagIllinois, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $120k/year
ApplyView job
NVIDIA3 hours ago

Service Reliability Engineer

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$168k – $333.5k/year
ApplyView job
Entarian3 hours ago

DevSecOps Engineer – Mid

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
BeyondTrust3 hours ago

Senior DevOps Engineer

CA flagCanada OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ontrac Solutions3 hours ago

Site Reliability Engineer

PK flagPakistan OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers