Principal Site Reliability Engineer – Lead

Posted 12 hours ago

This is a fully remote position, open to applicants in Poland.

📋 Description

• Contribute to the design, development, and management of a Golden Path platform utilizing Kubernetes and cloud-native technologies.

• Offer technical and architectural guidance for the Next Generation Control Plane.

• Drive the transition to a highly available, resilient, and distributed microservices architecture.

• Provide leadership, support, and mentoring to team members.

• Develop and maintain Service Level Objectives (SLOs) and Key Performance Indicators (KPIs).

• Collaborate with Engineering, Product, and Support teams.

• Participate in on-call duty rotations.

• Assist in the restoration and resolution of service-impacting incidents.

• Operate in an environment that prioritizes innovative solutions, rapid development cycles, and open communication.


⛳️ Requirements

• A minimum of 10 years of relevant professional experience.

• A Bachelor's degree in Computer Science or a related field, or equivalent experience.

• Extensive experience in building and managing highly available, fault-tolerant, and scalable production services using Kubernetes and other cloud-native technologies.

• Strong understanding of Linux internals, particularly in containerization and networking.

• A code-first approach to infrastructure operation, including automation practices.

• Proficiency in Go programming.

• Comfortable with using Python or Bash for scripting tasks.

• Familiarity with infrastructure-as-code tools like Crossplane, Pulumi, Terraform, or Ansible.

• Significant experience with modern observability tools such as OpenTelemetry, Prometheus, Grafana, or Loki.

• Capability to approach complex challenges with systematic curiosity.


🏝️ Benefits

• Comprehensive benefits supporting health, wellness, financial stability, and work-life balance.

• FlexBase flexible work arrangements: the option to work from home, in an office, or a mix of both.

People also viewed

Horizon3.ai11 hours ago

Staff Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$199.8k – $270k/year
ApplyView job
CLOUD MANTA GmbH11 hours ago

Senior DevOps Engineer, Containers & Private Cloud

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€70k – €80k/year
ApplyView job
Stefanini LATAM11 hours ago

Senior DevOps

AR flagArgentina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
PingWind Inc. (SDVOSB)12 hours ago

DevSecOps Engineer

US flagAlabama, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ad Hoc LLC12 hours ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $150k/year
ApplyView job
DraftKings Inc.13 hours ago

Database Reliability Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$112k – $140k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers