
Senior Manager, Site Reliability Engineer – FedRAMP, AWS GovCloud
Posted 11 hours ago

Posted 11 hours ago
This is a fully remote position, open to applicants in United States.
• Spearhead a newly formed cross-product SRE team dedicated to managing AWS GovCloud environments that support BeyondTrust products transitioning into FedRAMP authorization.
• Take ownership of reliability, change control, continuous compliance, and collaborative operational tasks for the designated environments.
• Design and maintain AWS GovCloud landing zones, which include multi-account frameworks, SCPs, guardrails, logging, and security tools.
• Manage the automated FedRAMP 20x continuous-compliance evidence pipeline along with monthly monitoring responsibilities.
• Oversee CI/CD pipelines, artifact management, security and quality gates, and ensure visibility into cloud costs.
• Regulate production changes while ensuring an auditable separation between development and production tasks.
• Handle release, patch, update, incident management, on-call duties, and post-incident processes.
• Facilitate automated transitions from development to staging and production.
• Organize the intake and prioritization of operational requests from feature teams and leadership.
• Lead a team of DevOps, SRE, and security automation engineers; establish accountability, delivery rhythms, planning cycles, and team health metrics.
• Collaborate with the VP of Application Engineering on workforce planning, hiring, role design, and onboarding processes.
• Possess staff-level technical expertise and the capacity to critically assess Terraform, pipeline definitions, and architectural proposals.
• Hands-on experience in developing and managing reliability and compliance systems.
• Proven people-management skills in hiring, developing, and leading a small SRE or DevOps team.
• Familiarity with AWS GovCloud environments and FedRAMP authorization/compliance.
• Knowledge of FedRAMP 20x Key Security Indicators and standard schema.
• Experience with multi-account AWS landing zones, SCPs, guardrails, logging, and security tools.
• Proficient in infrastructure-as-code provisioning.
• Experience managing CI/CD pipelines, build signing, artifact management, quality and security gates.
• Knowledge of production change governance, access controls, approval standards, and separation of responsibilities.
• Skilled in managing cloud fleet release and update processes, staged rollouts, golden images, and fleet upgrade orchestration at scale.
• Background in incident management, on-call design, and conducting post-incident reviews.
• Experience in leading teams of DevOps, SRE, and security automation engineers.
• Capable of managing intake, prioritization, team capacity, workforce planning, hiring, role design, and onboarding activities.
• A culture that promotes flexibility, trust, and ongoing learning.
• Comprehensive employee support and care initiatives.
• A workplace culture that emphasizes diversity and inclusion.
Entarian
BeyondTrust
Scribe
Hypertegrity AG
Get handpicked remote jobs straight to your inbox weekly.