Remotery

Database Reliability Engineer

Posted Jul 29

This is a fully remote position, open to applicants in Poland, +5 more states.

📋 Description

• Take charge of PostgreSQL production reliability, including HA design, Patroni, PgBouncer, replication, failover, upgrades, vacuum/bloat control, query tuning, locks, indexes, capacity, backups, PITR, and restore validation.

• Enhance disaster recovery and operational documentation through tested restores, mapped recovery paths, quantifiable RTO/RPO targets, runbooks, and secure maintenance strategies.

• Provide support for the broader database environment, including ClickHouse, MongoDB, and Redis.

• Resolve incidents, assess access and data-safety modifications, enhance monitoring, and familiarize yourself with existing production ClickHouse patterns.

• Streamline DBA workflows using Ansible, Terraform/OpenTofu, GitLab CI/CD, scripts, and create reproducible runbooks for provisioning, granting access, backups, restores, health checks, and ownership metadata.

• Assist in developing DBaaS-style self-service features so that engineering teams can request databases, access, credentials, and operational checks with reduced manual DBA involvement.

• Advance observability and incident response through Grafana, metrics, logs, SLOs, alert rules, Opsgenie routing, and effective communication during production challenges.


⛳️ Requirements

• Extensive hands-on experience with PostgreSQL in mission-critical production settings, usually 5+ years or equivalent expertise.

• Solid understanding of PostgreSQL internals and operations, including MVCC, WAL, transactions, locks, indexes, query planning, replication, autovacuum, bloat, major upgrades, backups, PITR, and restore validation.

• Demonstrated experience with highly available databases, with the ability to assess quorum, split-brain risks, failover, rollback, and recovery processes.

• Strong foundational knowledge of Linux and infrastructure, covering systemd, networking, storage, filesystems, CPU/memory/disk bottlenecks, TLS, DNS, firewalls, and root-cause analysis.

• Proficiency in automation with Ansible and scripting, along with strong advantages for experience with Terraform/OpenTofu, GitLab CI/CD, and merge-request based delivery.

• Capability to support multiple database engines. While you don’t need to be a ClickHouse expert initially, a willingness to learn quickly and take ownership is essential.

• Practical experience with AI engineering assistants like Claude and Codex, with an expectation to leverage them for enhancing speed and quality while personally validating generated SQL, commands, scripts, and operational insights.

• Proficient in English at an upper-intermediate level or higher to ensure effective communication of progress within teams.

• Nice to Have:

• Experience in ClickHouse operations, including replication, Keeper/ZooKeeper, MergeTree engines, distributed DDL, grants, row policies, backups, query troubleshooting, and cluster recovery.

• Knowledge of MongoDB replica sets and Percona Backup for MongoDB.

• Familiarity with Redis/Sentinel and understanding broker/cache failure modes.

• Insight into database observability, SLOs, golden signals, alert tuning, and actionable incident runbooks.

• Experience in building internal platforms, self-service portals, or DBaaS workflows for engineering teams.


🏝️ Benefits

• Emphasis on professional growth.

• Engaging and challenging projects.

• Fully remote position with flexible hours, allowing you to plan your day and work from any location globally.

• Paid 24 days of vacation annually, 10 national holidays, and unlimited sick leave.

• Coverage for private medical insurance.

• Reimbursement for co-working and gym/sports expenses.

• Budget allocated for educational pursuits.

• Opportunity to earn a reward for the most innovative idea that can be patented by the company.

People also viewed

CWILL19 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3720 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT20 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group20 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo21 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch21 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers