Remotery

Senior Staff DevOps Engineer

Posted Aug 4

This is a fully remote position, open to applicants in United States, +1 more state.

📋 Description

• Design, construct, and manage SailPoint’s global Identity Security Cloud infrastructure on AWS.

• Oversee the complete lifecycle of service mesh implementation across numerous microservices and production EKS clusters.

• Develop organization-wide service mesh standards, governance, onboarding runbooks, sidecar injection policies, traffic policy templates, and failure-mode playbooks.

• Mentor and enhance the skills of engineers in service mesh architecture, troubleshooting, performance tuning, and capacity planning.

• Design, manage, and optimize production Kubernetes clusters on AWS EKS, covering architecture, upgrades, node management, networking, storage, and multi-tenant isolation.

• Establish and promote Kubernetes standards and best practices across teams.

• Design and scale infrastructure to meet customer demand, data sovereignty requirements, and regional expansion goals.

• Automate deployment, monitoring, incident response, and capacity management using GitOps and CI/CD methodologies.

• Develop and refine operational practices, runbooks, and platform engineering standards.

• Collaborate with development teams to safely deploy new features and services into production.

• Support SailPoint’s PCI compliance initiatives through secure platform design, controls implementation, and audit readiness.

• Engage in and enhance the global follow-the-sun on-call rotation.

• Lead post-incident reviews and implement systemic fixes.

• Influence architectural direction, mentor engineers, and drive operational excellence without direct people management.


⛳️ Requirements

• Bachelor's and/or Master's degree in Computer Science or equivalent technical experience.

• 3+ years of direct experience in designing, implementing, and operating a service mesh at scale within production Kubernetes environments.

• 10+ years of experience in 24x7 production operations for highly available SaaS or cloud service environments.

• 10+ years of experience with containerization, virtualization, and configuration management technologies.

• 5+ years of hands-on experience with Kubernetes in large-scale production settings.

• 5+ years of experience using Terraform to manage infrastructure across various AWS accounts and regions.

• 5+ years of experience designing and implementing CI/CD pipelines, particularly for Terraform, Kubernetes, and microservices.

• 5+ years of experience with scripting/programming languages such as Python or Go.

• Strong proficiency in shell scripting.

• Comprehensive understanding of Linux, networking, distributed systems, and production troubleshooting.

• Proven experience in scaling a service mesh across a large fleet of microservices, including phased adoption strategies, sidecar resource management, control plane scaling, and performance tuning.

• Experience with monitoring and logging stacks such as Prometheus, Grafana, and OpenSearch or similar tools.

• Previous experience as a technical lead or Staff+ individual contributor within a global engineering organization.

• Strong interpersonal and teamwork skills, with the ability to establish and enforce processes and influence engineers across teams and different geographies.

• Ability to work effectively in an agile, entrepreneurial environment with global stakeholders.

• Candidates must be eligible to work in the United States or Canada.

• Availability to collaborate with global teams across US, EMEA, and APAC time zones.

• Participation in an on-call rotation is mandatory.

• Preferred: experience with multi-cluster or multi-region service mesh federation in production.

• Preferred: familiarity with PCI DSS and FedRAMP-adjacent practices.


🏝️ Benefits

• Medical, dental, and vision insurance.

• Short-term and long-term disability coverage.

• Life insurance and Accidental Death & Dismemberment (AD&D) coverage.

• Supplemental life insurance available for employees, spouses, and children.

• Flexible spending accounts for health care and dependent care.

• Limited purpose flexible spending account.

• 401(k) Savings and Investment Plan with company matching.

• Flexible vacation policy.

• 8 paid holidays annually.

• Sick leave.

• Paid parental leave.

• Employee Assistance Program (EAP) and access to Care Counselors.

• Options for Voluntary Legal Assistance, Critical Illness, Accident, Hospital Indemnity, and Pet Insurance.

• Health Savings Account (HSA) with employer contributions.

• Eligibility for the SailPoint Corporate Bonus Plan or role-specific commission may apply.

• Potential eligibility for equity participation.

• Approximately 10% travel flexibility.

People also viewed

TEKsystems23 hours ago

SRE – CloudOps, Practice Architect II

US flagIllinois OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
TEKsystems23 hours ago

SRE CloudOps Practice Architect II

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data1 day ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Level Data1 day ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data1 day ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data1 day ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers