Remotery

Principal Site Reliability Engineer

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Oversee project initiatives that involve the development and maintenance of platform features related to product reliability, networking, and cloud infrastructure.

• Commit to infrastructure-as-code on a daily basis.

• Occasionally develop and test application code using Python, Golang, and JavaScript.

• Guide and inspire service owners in deploying, measuring, monitoring, and operating services at scale.

• Engage in a weekly support rotation, which includes on-call pager responsibilities and providing support during working hours for platform users.

• Lead efforts in incident response, triage, and root cause analysis support.

• Collaborate with architects and product stakeholders on initiatives aimed at enhancing reliability.

• Pair-program and mentor junior Site Reliability Engineers (SREs).

• Manage a shared backlog, engage in weekly pair programming, perform peer reviews of work, and take part in blame-free retrospectives.


⛳️ Requirements

• Extensive experience in operating Kubernetes within highly distributed environments.

• Proven experience managing systems in Google Cloud Platform (GCP) or Amazon Web Services (AWS).

• Familiarity with monitoring, observability infrastructure, and industry best practices.

• Solid understanding of infrastructure-as-code methodologies, tools, and patterns.

• Some background in software development within Linux environments, ideally with Python and/or Golang.

• A minimum of six years of experience in systems operations or development.

• Customer-focused approach when supporting platform users and fostering organizational trust.

• A collaborative attitude for effective teamwork across various departments.


🏝️ Benefits

• Eligibility for bonuses.

• Stock equity options.

• Unlimited paid time off (PTO).

• Flexible work location options.

• Up to 24 weeks of parental leave.

• Comprehensive health benefits.

• Reasonable accommodations for disabilities during the hiring process, work, and access to benefits.

People also viewed

Level Data3 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Level Data3 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data3 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data3 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Identity Digital Inc.7 hours ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$175k – $220k/year
ApplyView job
Identity Digital Inc.7 hours ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$175k – $220k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers