Remotery

Senior Site Reliability Engineer

Posted 3 hours ago

This is a fully remote position, open to applicants in Brazil.

📋 Description

• Assist in the implementation of observability tools (Datadog and/or Azure Monitor/App Insights) while helping to establish SLO definitions, alert rules, and synthetic checks.

• Engage in a PagerDuty on-call rotation, which includes handling escalations and documenting incidents.

• Develop and maintain operational runbooks for incident response, rollback, and recovery procedures.

• Contribute to deployment automation efforts (using blue/green or canary patterns) and Infrastructure as Code.

• Work within Azure SQL and Cosmos DB environments, aiding in performance and cost optimization projects.

• Collaborate closely with engineers based in the US during overlapping working hours.


⛳️ Requirements

• Over 5 years of experience in SRE, DevOps, or cloud infrastructure positions.

• Strong practical knowledge of Microsoft Azure (including Azure SQL, Cosmos DB, Container Apps, App Service).

• Familiarity with observability tools (such as Datadog, Azure Monitor, or similar) and on-call/incident response processes.

• Understanding of Infrastructure as Code, with a preference for Terraform.

• Excellent written and verbal communication skills in English; daily interaction with US-based team members and occasionally with client stakeholders is expected.

• **Availability for significant overlap with US Eastern or Mountain time zones.**

• Experience in HIPAA-regulated settings, including managing PHI under a Business Associate Agreement (BAA) and adhering to least-privilege, audited access controls.

• Willingness to undergo a healthcare-industry-standard background check before gaining production access.

• **On-Call Expectations:**

• This position necessitates participation in a pager-based on-call rotation through PagerDuty, addressing SEV-1/SEV-2 incidents as part of a shared schedule with the SRE team. This is a fundamental aspect of the role, not a rare request.


🏝️ Benefits

• Work remotely.

• Vacation: 10 business days each year.

• Holidays: 5 National Holidays per year.

• Company Holidays: 5 Company Holidays annually (Christmas Eve, Christmas Day, New Year's Eve, New Year's Day, Zipdev Day).

• Parental Leave.

• Health Care Reimbursement.

• Active Lifestyle Reimbursement.

• Quarterly Home Office Reimbursement.

• Payroll Deduction Purchase Plans.

• Longevity Bonus.

• Continuous Learning Bonus.

• Access to Training and Professional Development Platforms.

• Did we mention it's REMOTE?!!

People also viewed

Fundraise Up3 hours ago

Senior DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€6,000 – €6,800/month
ApplyView job
Empower3 hours ago

Lead Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$114k – $165.3k/year
ApplyView job
Harrods3 hours ago

DevOps Manager

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Aufinity Group | España3 hours ago

Software Developer – DevOps

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Valtech4 hours ago

Senior Site Reliability Engineer

MK flagMacedonia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
One Identity4 hours ago

Staff Software Engineer – Reliability & Platform

GB flagUnited Kingdom OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers