
Engineering Manager, Reliability Engineering β EDA Infrastructure
Posted 10 hours ago

Posted 10 hours ago
This is a fully remote position, open to applicants in California, +3 more states.
β’ Lead a team and take ownership of the roadmap for operational processes and platforms, overseeing everything from requirements and delivery to adoption and outcomes.
β’ Establish the technical direction, prioritize tasks, and guide execution across engineering and operational domains.
β’ Collaborate with infrastructure, product, and security teams to create uniform practices for incident response, maintenance, on-call duties, issue management, and customer service readiness.
β’ Recruit and nurture engineers and technical leads, forming a team with clear ownership and accountability.
β’ Align priorities among teams, communicate progress and risks, and provide technical guidance during significant incidents.
β’ Enhance reliability and minimize manual tasks through automation, AI, and insights gained from operational events.
β’ Develop and maintain systems that support chip development.
β’ Bachelor's degree or equivalent experience.
β’ Over 10 years of experience in software engineering or a related field.
β’ More than 5 years of engineering leadership experience managing teams or complex technical programs.
β’ Familiarity with operational processes and the platforms that support them, including roadmap development, delivery, adoption, and enhancement.
β’ Strong technical judgment in software architecture, platform integration, and engineering trade-offs.
β’ Excellent communication skills with engineers, cross-functional partners, and executive stakeholders.
β’ Proven track record in developing engineers, expanding teams, and achieving results under pressure.
β’ Established standards for service readiness, including service ownership, support coverage, and reliability goals.
β’ Experience in building, integrating, and scaling platforms related to incident management, maintenance, customer experience management, on-call duties, and production readiness.
β’ Experience utilizing AI or LLMs to enhance triage, knowledge retrieval, incident analysis, or automation.
β’ Background in supporting EDA, large-scale computing, or hybrid infrastructure with intricate dependencies and stringent availability requirements.
β’ Equity.
β’ Comprehensive benefits package.
Probis
Entarian
Velera
RR Donnelley
Get handpicked remote jobs straight to your inbox weekly.