SRE Engineer

Posted 8 hours ago

This is a fully remote position, open to applicants in Brazil.

📋 Description

• Join Verity as an SRE Engineer, contributing to a digital transformation and engineering consultancy.

• Establish and monitor reliability metrics.

• Deploy observability, monitoring, alerting systems, and Application Performance Management (APM).

• Oversee latency, traffic, errors, saturation, availability, and performance metrics.

• Proactively prevent, identify, and resolve incidents.

• Perform root cause analysis and implement preventive measures.

• Detect risks, bottlenecks, and single points of failure.

• Aid in the design of resilient, scalable, and highly available solutions.

• Automate operational processes to minimize manual tasks.

• Manage and enhance Kubernetes and Docker environments.

• Assist in capacity planning, business continuity, and disaster recovery strategies.

• Engage in deployments and support application stabilization.

• Collaborate with teams to enhance reliability from the design phase onward.

• Develop and maintain dashboards, alerts, procedures, and operational documentation.


⛳️ Requirements

• Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), Service Level Agreements (SLAs), Mean Time to Recovery (MTTR), and Mean Time to Detect (MTTD).

• Implement observability, monitoring, alerting systems, and APM.

• Track latency, traffic, errors, saturation, availability, and performance.

• Prevent, identify, and resolve incidents effectively.

• Conduct root cause analyses and establish actions to avert recurrence.

• Recognize risks, bottlenecks, and single points of failure.

• Assist in designing resilient, scalable, and highly available solutions.

• Automate operational activities to lessen manual tasks.

• Operate and advance Kubernetes and Docker environments.

• Support capacity planning, business continuity, and disaster recovery efforts.

• Participate in deployments and aid in application stabilization.

• Collaborate with teams to enhance reliability from the design stage onward.

• Create and maintain dashboards, alerts, procedures, and operational documentation.

• Possess hands-on experience with SRE and reliability metrics, including SLI, SLO, SLA, MTTR, and MTTD (as inferred from a required application question).


🏝️ Benefits

• Meal allowance

• Food allowance

• Work-from-home allowance

• Medical insurance

• Dental insurance

• Life insurance

• Birthday day off

• Total Pass / Wellhub

• Boon Saúde app

• Discount partnerships

• Discounts at participating establishments and educational institutions

• Welcome kit

• Onboarding

• Verity Learning

• Verity Break

• #VerityComVocê

• Access to professional development courses

People also viewed

Fundrise7 hours ago

AI Infrastructure Deployment Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$200k – $230k/year
ApplyView job
Méliuz8 hours ago

Senior SRE Analyst

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ODILO8 hours ago

DevOps Engineer

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Comet10 hours ago

Deployment Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$150k – $200k/year
ApplyView job
Element 8411 hours ago

Senior DevOps Engineer, NOAA Badge Required

US flagArizona, +20 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$145k – $180k/year
ApplyView job
CONVOTIS12 hours ago

Infrastructure Automation Engineer – Ansible, AWX, DevOps

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers