Remotery

Principal Operations Engineer, Reliability

Posted Jul 28

This is a fully remote position, open to applicants in United States.

📋 Description

• Take charge of fleet reliability engineering: establish availability targets, assess them transparently, and bridge any gaps.

• Conduct root cause analyses on the fleet's most significant incidents and implement corrective actions throughout all sites.

• Develop a failure data pipeline for both facilities and hardware that transforms incident history into engineering priorities.

• Formulate the maintenance strategy (reliability-centered, condition-based) to ensure the fleet concentrates efforts where failure data indicates.


⛳️ Requirements

• You have taken responsibility for the reliability of critical infrastructure and successfully influenced the availability metrics, not just reported them.

• You have spearheaded root cause analyses that identified the genuine cause rather than the more convenient one.

• You are proficient in working with failure data: Weibull, Pareto, and FMEA are tools you actively utilize, not just terminology you recognize.

• You ensure corrective actions are completed across teams under your influence, even if you don't manage them directly.

• Bonus: Experience in data center or power generation reliability. Knowledge of liquid cooling systems. Familiarity with CMMS analytics. Possession of CRE or CMRP certification.


🏝️ Benefits

• Competitive base salary

• Equity offered for all full-time positions

• Comprehensive benefits package

• Commission plans available if applicable

People also viewed

DATAGROUP2 days ago

DevOps Engineer

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ambush2 days ago

DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
DuoKey2 days ago

DevOps Engineer

MU flagMauritius OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
TEKsystems3 days ago

SRE – CloudOps, Practice Architect II

US flagIllinois OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
TEKsystems3 days ago

SRE CloudOps Practice Architect II

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data3 days ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers