Remotery

Staff Site Reliability Engineer

Posted Aug 1

This is a fully remote position, open to applicants in United States.

📋 Description

• As a Staff Site Reliability Engineer at Filevine, you will serve as the senior technical authority on the SRE team and act as a strategic partner to engineering leadership.

• Your role extends beyond system maintenance; you will influence engineering culture, establish the technical standards for how Filevine operates in production, and connect high-level business objectives with effective, internet-scale technical execution.

• You will take ownership of the roadmap for two pivotal SRE domains — Observability & Alerting and Platform Infrastructure — ensuring the team addresses reliability challenges permanently rather than merely managing them as ongoing issues.

• You will function as the senior individual contributor counterpart to the Engineering Manager, taking responsibility for technical accuracy.

• You will collaborate with the Reliability Architect and engineering leadership on major technical decisions, mentor engineers of varying experience levels, and impact the reliability strategy throughout the organization.

• At Filevine, reliability is crucial to protecting revenue.

• As the senior technical voice, you will ensure that uptime, incident responses, and every change in production adhere to the operational standards required by the business.


⛳️ Requirements

• A minimum of 12 years of experience in software engineering, infrastructure, platform engineering, or SRE, with at least 6 years in SRE and 3 years leading intricate, cross-functional technical initiatives for distributed production systems.

• Expert-level knowledge in observability and platform infrastructure, along with extensive skills in incident response, capacity planning, automation, and reliability engineering.

• Advanced proficiency with a major container-orchestration platform, ideally Kubernetes, and familiarity with an observability platform such as New Relic, Datadog, or similar.

• Strong software engineering skills in Python, Go, Bash, or another general-purpose programming language, with experience in developing production tooling, automation, or platform capabilities.

• Demonstrated ability to mentor engineers and effectively convey technical risks to engineering, product, and executive stakeholders.

• Experience in a regulated environment such as FedRAMP, CJIS, HIPAA, SOC 2, or PCI is highly preferred.


🏝️ Benefits

• Medical, Dental, & Vision Insurance (for full-time employees)

• Competitive & Fair Pay

• Maternity & paternity leave (for full-time employees)

• Short & long-term disability

• Opportunity to learn from a dedicated leadership team

• Top-of-the-line company swag

People also viewed

DATAGROUP2 days ago

DevOps Engineer

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ambush2 days ago

DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
DuoKey2 days ago

DevOps Engineer

MU flagMauritius OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
TEKsystems3 days ago

SRE – CloudOps, Practice Architect II

US flagIllinois OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
TEKsystems3 days ago

SRE CloudOps Practice Architect II

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data3 days ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers