Senior Database Reliability Engineer

Posted Aug 28

This is a fully remote position, open to applicants in Brazil.

📋 Description

• Ensure the continuous availability, performance, and dependability of MySQL and PostgreSQL databases as essential services from an SRE/DBRE viewpoint.

• Establish, execute, and consistently enhance SLOs, SLIs, and SLAs, while monitoring error budgets and taking proactive measures.

• Design and manage scalable, resilient cloud architectures, including Multi-AZ, read replicas, sharding, and partitioning.

• Collaborate with development teams to optimize query, index, execution plans, parameters, connections, and connection pooling.

• Define and uphold strategies for backup, restoration, point-in-time recovery, and disaster recovery.

• Automate operational tasks using Infrastructure as Code (IaC) and scripting to minimize manual processes and reduce toil.

• Strategize for capacity and data growth while optimizing cloud resources and controlling costs.

• Implement observability through metrics, logs, traces, dashboards, and alerts.

• Plan and carry out engine and workload migrations and upgrades with minimal disruption.

• Establish and advocate for standards, guidelines, and frameworks concerning data modeling, access, and operations.

• Evaluate and resolve critical incidents, leading root cause analyses and implementing preventive measures.

• Ensure security and compliance, focusing on encryption, access controls, secrets management, data masking, LGPD, and PCI DSS.

• Conduct design reviews and schema assessments.

• Advocate for operational excellence, reliability, and best practices in data engineering.

• Assess new AWS data technologies and services.


⛳️ Requirements

• Demonstrated professional experience in administering, operating, and tuning relational databases within large-scale production environments.

• Extensive knowledge of MySQL, including replication, storage engines, query optimization, parameters, and troubleshooting.

• Comprehensive understanding of PostgreSQL, covering MVCC, vacuum/autovacuum, indexes, execution plans, extensions, and replication.

• Experience with AWS and its data, networking, and security services: RDS, Aurora, DMS, S3, IAM, VPC, KMS, Secrets Manager, and Parameter Store.

• Background in applying SLOs/SLIs, error budgets, observability, and toil reduction to data services.

• Proficiency in tuning and optimizing performance, including queries, indexes, locks, contention, and connection pooling.

• Familiarity with PgBouncer, ProxySQL, or RDS Proxy.

• Experience with high availability, replication, failover, backup, restore, and disaster recovery strategies.

• Knowledge of IaC and automation using Terraform or AWS CDK (Python).

• Skills in Shell/Bash scripting.

• Experience with monitoring tools such as Grafana, Prometheus, Datadog, New Relic, CloudWatch, or Performance Insights.

• Understanding of data security and compliance, focusing on encryption, access management, secrets management, and auditing.

• Familiarity with CI/CD practices and schema versioning/migrations using tools like Flyway or Liquibase.

• Capability to create and maintain technical documentation, runbooks, and diagrams.

• Preferred qualifications include AWS certifications; experience in fintech or regulated environments; large-scale migrations; NoSQL/in-memory databases; streaming; FinOps; containers; Chaos Engineering; advanced observability; networking; Git/GitHub/GitFlow; and proficiency in Python, Go, Java, or Node.js.


🏝️ Benefits

• Comprehensive medical and dental insurance with no co-pay.

• Life insurance coverage.

• Allowance for prescription medications.

• Fitness allowance for health-related activities.

• Four complimentary therapy or nutritionist sessions each month.

• Quick massage services available at headquarters.

• Flexible meal benefit provided on a Visa card.

• Free food available at headquarters.

• Childcare allowance to support working parents.

• Parental support program for new parents.

• Extended maternity and paternity leave policies.

• Access to an in-house training platform.

• Education allowance that covers 70% of tuition for undergraduate and language courses, as well as professional courses and books.

• Home office allowance to support remote work.

• Provision of necessary work equipment.

• Furniture allowance for home office setup.

• Partnership with WOBA for coworking space access throughout Brazil.

• Day off during your birthday month.

• Happy hour allowance for team bonding.

• Referral bonus for successfully bringing new hires.

• Annual performance-based bonus structure.

• Stock options plan for employees.

• No dress code policy to promote comfort at work.

People also viewed

CI&T17 hours ago

Senior/Specialist DevOps, Azure

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sprezzatura2 days ago

Salesforce Release Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $150k/year
ApplyView job
Clinician Nexus2 days ago

Manager, DevOps

US flagArizona, +13 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$133.1k – $221.9k/year
ApplyView job
Lenovo2 days ago

CI/CD Engineer

US flagNorth Carolina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$146.2k – $224.1k/year
ApplyView job
MRSOOL | مرسول2 days ago

Site Reliability Engineer II

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
CEQUENS2 days ago

DevOps Engineer

EG flagEgypt OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers