
Senior Director, Cloud Operations
Posted 5 hours ago

Posted 5 hours ago
This is a fully remote position, open to applicants in United States.
• Take charge of comprehensive cloud operations for Foundant's SaaS product suite.
• Ensure the availability, performance, and dependability of production systems.
• Spearhead availability incident management, including triage, resolution, and root cause analysis.
• Deliver incident reports and root cause analyses to both internal and external stakeholders.
• Supervise configuration and change management across automated CI/CD and manually managed environments.
• Oversee infrastructure hosting expenses across the US, Canada, EU, Australia, and UK regions.
• Build and lead a team of system administrators and Site Reliability Engineers (SREs).
• Mentor team members towards achieving technical excellence and accountability for operational results.
• Collaborate with Development, Technical Security, and Support teams for operational readiness, incident response, and handoffs.
• Establish and continuously enhance monitoring, alerting, and performance management methodologies.
• Translate Foundant's strategic vision into a cloud operations approach across five products.
• Set and enhance availability and performance benchmarks using data and monitoring insights.
• Make prompt decisions during live availability incidents.
• Clearly communicate root causes of incidents and operational modifications.
• Develop team members through feedback, technical coaching, and opportunities for ownership.
• Perform additional duties as assigned.
• Over 10 years of progressive experience in cloud operations or infrastructure leadership.
• More than 5 years in a director-level position.
• Extensive, hands-on expertise in operating and troubleshooting multi-cloud environments, including AWS and Azure.
• Experience in managing availability and performance of SaaS products deployed across various geographic regions.
• Familiarity with regional data residency and compliance issues in the US, Canada, EU, UK, and Australia.
• Proven experience leading availability incident management and root cause analysis for production environments.
• Experience overseeing configuration and change management within infrastructure-as-code/CI-CD and traditional manually managed release environments.
• Skilled in managing cloud infrastructure hosting costs at scale, including forecasting and cost optimization.
• Experience in building, managing, and developing teams of system administrators and/or Site Reliability Engineers.
• Bachelor's degree in Computer Science, Information Systems, or a related field is preferred.
• Must have legal eligibility to work in the United States.
• Competitive salary.
• Comprehensive benefits package.
• Tuition reimbursement program.
• Lifestyle reimbursement options.
• Customized mindfulness initiatives.
• Tailored fitness programs.
• Flexible Paid Time Off (PTO) policy.
• Opportunities for professional and personal development.
• Collaboration across teams with exposure to diverse ideas, expertise, and projects.
• Opportunities for internal mobility and career advancement.
• Autonomy and responsibility in your role.
• Employee recognition initiatives.
• Accommodations provided throughout the interview and employment process.
Health Care Service Corporation
Health Care Service Corporation
inquirED
Get handpicked remote jobs straight to your inbox weekly.