
Staff Site Reliability Engineer
Posted Aug 1

Posted Aug 1
This is a fully remote position, open to applicants in United States.
• As a Staff Site Reliability Engineer at Filevine, you will serve as the senior technical authority on the SRE team and act as a strategic partner to engineering leadership.
• Your role extends beyond system maintenance; you will influence engineering culture, establish the technical standards for how Filevine operates in production, and connect high-level business objectives with effective, internet-scale technical execution.
• You will take ownership of the roadmap for two pivotal SRE domains — Observability & Alerting and Platform Infrastructure — ensuring the team addresses reliability challenges permanently rather than merely managing them as ongoing issues.
• You will function as the senior individual contributor counterpart to the Engineering Manager, taking responsibility for technical accuracy.
• You will collaborate with the Reliability Architect and engineering leadership on major technical decisions, mentor engineers of varying experience levels, and impact the reliability strategy throughout the organization.
• At Filevine, reliability is crucial to protecting revenue.
• As the senior technical voice, you will ensure that uptime, incident responses, and every change in production adhere to the operational standards required by the business.
• A minimum of 12 years of experience in software engineering, infrastructure, platform engineering, or SRE, with at least 6 years in SRE and 3 years leading intricate, cross-functional technical initiatives for distributed production systems.
• Expert-level knowledge in observability and platform infrastructure, along with extensive skills in incident response, capacity planning, automation, and reliability engineering.
• Advanced proficiency with a major container-orchestration platform, ideally Kubernetes, and familiarity with an observability platform such as New Relic, Datadog, or similar.
• Strong software engineering skills in Python, Go, Bash, or another general-purpose programming language, with experience in developing production tooling, automation, or platform capabilities.
• Demonstrated ability to mentor engineers and effectively convey technical risks to engineering, product, and executive stakeholders.
• Experience in a regulated environment such as FedRAMP, CJIS, HIPAA, SOC 2, or PCI is highly preferred.
• Medical, Dental, & Vision Insurance (for full-time employees)
• Competitive & Fair Pay
• Maternity & paternity leave (for full-time employees)
• Short & long-term disability
• Opportunity to learn from a dedicated leadership team
• Top-of-the-line company swag
DATAGROUP
Ambush
DuoKey
TEKsystems
Get handpicked remote jobs straight to your inbox weekly.