Remotery

System Reliability Engineer

Posted Jul 17

This is a fully remote position, open to applicants in Costa Rica.

📋 Description

• Ensure continuous monitoring and seamless operation of the company's services.

• Develop and maintain alerting protocols and operational runbooks.

• Conduct triage of incoming incidents along with the initial diagnosis of issues.

• Establish and uphold escalation procedures for incident response.

• Execute technical incident resolution tasks as per the established runbooks.

• Engage in on-call rotations and facilitate post-incident analyses (RCA/postmortems).

• Strive for enhancements in observability coverage while minimizing alert noise and false positives.

• Work collaboratively with development and infrastructure teams to pinpoint reliability risks and implement proactive measures.


⛳️ Requirements

• Proficiency with observability tools such as Grafana, ELK, and VictoriaMetrics.

• Experience in working with Linux operating systems.

• Familiarity with Kubernetes (k8s) environments.

• Knowledge of AWS and Azure cloud platforms.

• Capability to analyze incidents, identify root causes, and suggest remediation steps.


🏝️ Benefits

• Comprehensive Health Coverage – Fully employer-paid medical, dental, and vision insurance for employees and eligible dependents, including virtual care services.

• Wellbeing & Mental Health Support – Access to an Employee Assistance Program (EAP) that includes confidential therapy sessions, as well as legal and financial counseling services.

• Financial Protection Benefits – Company-provided life and disability insurance designed to support employees and their families.

• Paid Time Off & Global Recharge Days – Vacation time, statutory holidays, and additional company-wide VeeaMe Days dedicated to rest, wellbeing, and self-care.

• Family-Friendly Leave Programs – Competitive maternity, paternity, adoption, and other leave benefits that support employees during significant life events.

• Give Back to Your Community – Employees are granted paid volunteer time each year through the Veeam Cares program.

People also viewed

CWILL11 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3712 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT13 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group13 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo14 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch14 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers