Remotery

Senior Lead Database Reliability Engineer

Posted 4 days ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Lead the technical roadmap focused on database reliability across PostgreSQL, MySQL, MongoDB, Redis, ScyllaDB, Aerospike, and managed cloud services, while designing architecture that ensures high availability, replication, partitioning, storage, and connection management.

• Create and implement automation-first database platforms by developing Kubernetes operators, utilizing infrastructure as code, employing GitOps workflows, and producing production-grade tooling in Go or Python to automate provisioning, failover, backups, schema migrations, and lifecycle management.

• Drive operational excellence by establishing service level objectives, monitoring the health, capacity, and performance of databases, resolving recurring reliability issues, validating backup and recovery processes, and leading critical production incidents to resolution while fostering continuous improvement.

• Enhance database performance and cost-effectiveness in both cloud and on-premises settings by overseeing capacity planning, resource efficiency, storage optimization, workload consolidation, and performance tuning for large-scale systems.

• Collaborate closely with application engineering teams to implement safe database practices, which include schema reviews, migration strategies, query optimization, connection management, and zero downtime deployment methods.

• Utilize AI to boost engineering productivity and database operations through intelligent observability, anomaly detection, root cause analysis, documentation, predictive insights, and evaluation of AI-generated code to ensure reliability and security.

• Mentor engineers across the organization by sharing best practices, leading design and code reviews, shaping technical direction, supporting recruitment efforts, and enhancing the overall maturity of database reliability engineering.


⛳️ Requirements

• A minimum of 6 years of experience in Database Reliability Engineering, Database Platform Engineering, or Site Reliability Engineering with a strong focus on databases.

• Extensive expertise in at least one major relational database, ideally PostgreSQL, alongside operational experience with technologies such as MySQL, MongoDB, Redis, ScyllaDB, Aerospike, Aurora, Cloud SQL, and other managed cloud database services.

• Significant experience in building and managing stateful workloads on Kubernetes using technologies such as StatefulSets, Persistent Volumes, database operators, Terraform, Pulumi, FluxCD, ArgoCD, GKE, and EKS.

• Hands-on software development experience with Go or Python to create automation, platform tooling, Kubernetes controllers, APIs, and infrastructure.

• A data-driven, automation-first mindset with a proven record of enhancing reliability through observability, monitoring, service level objectives, capacity planning, performance optimization, and self-service engineering solutions.

• Practical experience with AI tools such as Claude, GitHub Copilot, Cursor, MCP, or similar technologies to refine design, coding, documentation, troubleshooting, and operational workflows, while applying sound engineering judgment to validate AI-generated outputs.

• Outstanding leadership and communication skills, with a history of mentoring engineers, influencing architectural decisions, collaborating across engineering teams, producing clear technical documentation, and driving continuous improvement in highly available production environments.


🏝️ Benefits

• Plus bonus

• Equity

• Benefits as applicable

People also viewed

Fundraise Up12 hours ago

Senior DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€6,000 – €6,800/month
ApplyView job
Empower12 hours ago

Lead Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$114k – $165.3k/year
ApplyView job
Harrods12 hours ago

DevOps Manager

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Aufinity Group | España12 hours ago

Software Developer – DevOps

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Zipdev12 hours ago

Senior Site Reliability Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Valtech12 hours ago

Senior Site Reliability Engineer

MK flagMacedonia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers