Remotery

Senior Site Reliability Engineer

Posted Aug 5

This is a fully remote position, open to applicants in North America.

📋 Description

• Manage daily production operations, including on-call duties, incident management, post-incident reviews, and subsequent actions.

• Establish and enhance SLIs/SLOs and error budgets, enabling product teams to function effectively within these parameters.

• Improve observability across metrics, logs, traces, and alerting systems.

• Deploy cloud resources and Kubernetes workloads through code within a GitOps framework.

• Ensure the reliability of PostgreSQL by optimizing performance, reviewing schema and migrations, conducting online migrations on large tables, and managing HA/DR and CDC pipelines.

• Provide mentorship to engineers on reliability and database principles through code reviews, design discussions, and collaborative work.


⛳️ Requirements

• Over 4 years of experience in SRE, DevOps, Platform/Infrastructure, or backend engineering with substantial ownership of production operations.

• Practical experience in managing production services on Kubernetes and delivering infrastructure as code using a GitOps approach.

• Strong understanding of PostgreSQL in a production environment, including query plans, pg_stat_*, indexing, schema considerations, and safe online migrations.

• Knowledge of cloud networking essentials: VPCs, routing, L4/L7 load balancing, DNS, and TLS.

• Familiarity with a contemporary observability stack.

• Competent in Linux at an operational level.

• Experienced in incident response, structured debugging, and conducting post-incident analyses.

• At least conversational proficiency in Go or Python.

• Excellent written and verbal communication skills.

• A genuine passion for databases and a desire to develop PostgreSQL/DBA expertise.

• Preferred qualifications include advanced PostgreSQL knowledge, typed SQL access layers in Go, experience with large-scale messaging systems, security and compliance insights in regulated environments, and familiarity with trading, brokerage, or regulated fintech.


🏝️ Benefits

• Competitive salary and stock options.

• Comprehensive health benefits.

• New hire home-office setup: one-time allowance of USD $500.

• Monthly stipend of USD $150 via a Brex Card.

People also viewed

DATAGROUP2 days ago

DevOps Engineer

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ambush2 days ago

DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
DuoKey2 days ago

DevOps Engineer

MU flagMauritius OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
TEKsystems3 days ago

SRE – CloudOps, Practice Architect II

US flagIllinois OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
TEKsystems3 days ago

SRE CloudOps Practice Architect II

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data3 days ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers