
Manager, Platform Engineering
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in United States.
• Lead the platform engineering team: take ownership of its delivery, reliability commitments, overall health, and the professional growth and career development of its engineers, while maintaining a hands-on approach in infrastructure and tooling.
• Ensure that LegitScript’s infrastructure is auditable: manage environment separation, change control, access management, and audit logging, making sure that production changes are traceable and defensible, in alignment with the Secure SDLC Policy, Change Management Policy, and SOC 2 evidence requirements.
• Spearhead the transition to a standardized Kubernetes deployment path, executed incrementally at natural replacement points, while ensuring operational continuity for the dependent teams and services.
• Develop and maintain the developer tooling and self-service deployment pathways that enable engineers to deploy application code to production without manually managing infrastructure: creating paved roads, golden paths, and guardrails that facilitate a safe and easy process.
• Take ownership of the CI/CD pipeline and its associated controls, including branch protection, required peer review, automated security and quality testing, and gated deployment, ensuring that the pipeline enforces our engineering and security standards at scale.
• Establish infrastructure-as-code standards using Terraform, bringing manual or unmanaged production changes under version control and change management to eliminate drift that leads to technical debt and audit gaps.
• Minimize routine operational toil through automation and replatforming, allowing volume growth to shift away from being a headcount equation and enabling engineering capacity to focus on higher-value tasks.
• Take charge of reliability, monitoring, alerting, and incident response for infrastructure as essential concerns; act as an escalation point and coordinate responses for infrastructure incidents.
• Maintain the boundary with the application platform: provide the infrastructure substrate and tooling utilized by the application-platform and domain teams, collaborating with the Staff Engineer responsible for that platform rather than duplicating or absorbing their scope.
• Set and uphold standards for infrastructure code quality, testing, documentation, and review, elevating expectations for the engineers and partners working on the platform through clarity and example.
• Deliberately manage cloud costs and capacity, making explicit trade-offs between performance, reliability, and spending instead of leaving them incidental.
• Cultivate a culture of accountability, collaboration, and continuous improvement.
• Extensive, hands-on experience in platform and infrastructure engineering, particularly with AWS, Linux, and container-first workflows.
• Experience operating production Kubernetes environments: not only managing clusters but also providing other engineering teams with a safe, standardized deployment path.
• Strong background in infrastructure-as-code practices using Terraform, including managing existing or unmanaged infrastructure under version control and change management.
• Proven experience in the design and ownership of CI/CD pipelines, including the security, quality, and approval gates that uphold engineering standards on every change.
• A demonstrated history of building internal developer tooling and self-service deployment solutions that abstract infrastructure for application engineers.
• Familiarity with designing infrastructure and change processes to ensure auditability and compliance, including environment separation, access control, audit logging, and providing evidence for frameworks like SOC 2.
• Strong expertise in observability, monitoring, alerting, and incident response for infrastructure.
• Experience with authentication and authorization platforms such as Okta or Auth0.
• Proven ability to inherit unfamiliar, legacy, or poorly maintained infrastructure and enhance it significantly, demonstrating strong troubleshooting and root cause analysis skills.
• Experience in people leadership: managing, mentoring, and developing engineers while remaining technically engaged, and building consensus across teams to establish shared infrastructure and security practices.
• Demonstrated ability to influence without direct authority and to establish standards that are adopted by other engineers and teams.
• Comfortable navigating through change: leading an incremental migration effectively and establishing durable engineering practices from the ground up.
• Ability to clearly communicate technical concepts and trade-offs to non-technical stakeholders, including leadership.
• Proficient in speaking, reading, comprehending, and writing English.
• Competitive compensation.
• Flexible work options.
• A team dedicated to your success.
Yurts
First American
RoadRunner
Famedly GmbH
Get handpicked remote jobs straight to your inbox weekly.