
Senior Software Engineering Manager, Managed Gateways SREs
Posted Jul 23

Posted Jul 23
This is a fully remote position, open to applicants in Canada.
• Establish Kong's Managed Gateways SRE team in Toronto from the ground up, focusing on recruitment, setting performance standards, and being actively involved in enterprise implementations as the team develops.
• Serve as a key contributor to essential implementations and reliability initiatives during the initial phases, achieving outcomes that significantly enhance Managed Gateways' growth and business performance.
• Lead, mentor, and cultivate a high-performing team of Site Reliability Engineers committed to Kong's Managed Gateway offerings, directly supporting our crucial enterprise customer base across the Americas and Europe.
• Design and implement resilient, scalable, and fault-tolerant cloud-native systems utilizing technologies such as Kubernetes, Golang, and major cloud service providers.
• Manage the complete operational lifecycle, from proactive monitoring and alerting to incident response and blameless post-mortems, ensuring ongoing service improvements.
• Enhance developer satisfaction and operational efficiency through automation, self-service tools, and optimized workflows for deploying and managing API gateways, while actively mitigating technical debt and minimizing operational toil.
• Establish, monitor, and report on essential SLOs and SLIs, advocating for architectural best practices that maintain the performance and resilience of Managed Gateways as it expands.
• Collaborate across teams with Product, engineering, and support to influence roadmap decisions and guarantee operational preparedness for new features.
• Proven experience in leading and managing Site Reliability Engineering or DevOps teams within a fast-paced, high-growth setting.
• In-depth expertise in designing, deploying, and operating highly available distributed systems on cloud platforms (AWS, Azure, or GCP).
• Significant hands-on experience with Kubernetes (k8s) and container orchestration in live production environments.
• Proficiency in Golang or comparable modern programming languages for infrastructure automation and service development.
• Strong grasp of observability principles and familiarity with tools such as Prometheus, Grafana, OpenTelemetry, or similar.
• Proven ability to manage critical incidents, conduct root cause analysis, and implement effective preventive strategies.
• Knowledge of API gateway technologies, service mesh, or network proxies is a notable advantage.
• Health insurance
• 401(k) plan
• Short and long-term disability benefits
• Basic life and AD&D insurance
Remote People
Bestow
Virta Health
Space Inch
Get handpicked remote jobs straight to your inbox weekly.