Remotery

Staff Software Engineer – Databases SRE

Posted Jul 15

This is a fully remote position, open to applicants in Spain.

📋 Description

• Collaborate closely with product engineering teams (embedded model)

• Take ownership of production reliability for high-SLA and complex customer environments

• Design and implement automation solutions to enhance our reliability practices

• Ensure our customers achieve their SLO targets

• Define and refine per-tenant SLOs and reliability models

• Proactively minimize SLO burn to prevent recurring incidents

• Act as the primary escalation point and on-call for relevant incidents

• Lead incident response that affects customers and conduct post-incident reviews

• Contribute to design documentation and code evaluations

• Influence feature design to guarantee production scalability and operability

• Develop automation to eliminate unnecessary tasks where necessary

• Enhance alert quality and decrease excessive escalations


⛳️ Requirements

• Over 8 years of engineering experience, with at least 4 years in SRE/CRE/production engineering. Preference for candidates with formal customer reliability engineering experience.

• Extensive experience with Kubernetes in AWS, GCP, or Azure, along with familiarity with infrastructure-as-code tools (e.g., Helm, Terraform, Jsonnet, etc.).

• Strong experience in technical leadership, guiding a team through projects, mentoring fellow engineers, and acting as a force-multiplier.

• Experience managing multi-tenant systems in a production environment.

• Proven expertise in designing and implementing SLOs.

• Proficiency in one or more programming languages (e.g., Go, Python, Java, etc.).

• Knowledge of Linux operating systems internals, along with some understanding of networking, cloud storage, and scaling.

• Excellent problem-solving and troubleshooting capabilities.

• Experience in participating calmly and actively in blame-free Incident Response, following up on actions, and crafting high-quality Post Incident Reviews (PIRs).

• Ability to analyze performance, scalability, and failure modes.

• Comfortable working within an engineering team that encourages a strong sense of autonomy and self-direction.

• Capacity to collaborate deeply with product engineering teams.

• We highly value individuals who are intellectually curious, lean towards transparency, have a strong bias for action, and demonstrate kindness (this is essential!).


🏝️ Benefits

• 100% Remote, Global Culture

• Scaling Organization

• Transparent Communication

• Innovation-Driven

• Open Source Roots

• Empowered Teams

• Career Growth Pathways

• Approachable Leadership

• Passionate People

• In-Person onboarding

• Balance is Key

People also viewed

The CodestJul 26

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
IRIUMJul 26

Ingeniero/a Cloud DevOps

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€33k – €40k/year
ApplyView job
SólidesJul 26

Senior DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ResilincJul 25

Junior/Senior Site Reliability Engineer – Night Shift

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity GroupJul 25

Senior SRE / DevOps Engineer

Anywhere in the WorldFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
HOESSLER & HOESSLERJul 25

DevOps Software Engineer – Career Ambitions

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€65k – €75k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers