Remotery

Senior Site Reliability Engineer – Kubernetes

Posted Jun 29

This is a fully remote position, open to applicants in Poland.

📋 Description

• Manage and maintain Linux-based infrastructure (Debian/Ubuntu).

• Deploy, oversee, and scale Kubernetes clusters across bare-metal, virtualized, and on-premises environments.

• Supervise the complete lifecycle of clusters including upgrades, node pools, networking, storage, and security enhancements.

• Introduce automation for provisioning and operations utilizing Ansible, Bash/Python, and GitOps methodologies.

• Design and uphold networking architecture, covering VLANs, L2/L3 routing, VPNs, and multi-site connectivity.

• Create automated deployment workflows (PXE boot, Preseed, cloud-init).

• Deploy and sustain observability stacks (Prometheus/Grafana, Loki, ELK, Graylog).

• Spearhead incident response and escalation processes across the platform.

• Enhance system availability and decrease latency at all levels.

• Establish and implement SLOs/SLIs across various infrastructure levels (physical network/hardware, platform virtualization, software services).

• Streamline alerting and monitoring pipelines to yield actionable insights.

• Create and maintain on-call schedules to guarantee coverage across different timezones.

• Develop Standard Operating Procedures (SOPs) for consistent operations and maintenance tasks.

• Coordinate physical maintenance for Policlouds (periodic maintenance, hardware issues, DC-Ops).

• Oversee virtualization and orchestration layers (OpenStack, Proxmox, VMware).

• Assist in the development and maintenance of overall architecture across all products.

• Plan resources for future initiatives, considering demand and growth forecasts.

• Collaborate with development teams to enhance overall quality and optimize resource usage.

• Work in conjunction with cross-functional stakeholders (Hivenet, Policloud, Customer Success teams).


⛳️ Requirements

• Extensive, hands-on experience with operating Kubernetes in production settings.

• Strong network engineering expertise (VLANs, L2/L3 routing, VPNs, multi-site connectivity) is crucial for this position.

• Proficient in Linux systems administration (Debian/Ubuntu).

• Solid grasp of networking principles and the capability to design intricate network architectures.

• Experience in building and maintaining automation workflows (Ansible, Bash/Python, Git-based).

• Familiarity with observability stacks such as Prometheus, Grafana, ELK, Loki, or Graylog.

• Background in virtualization technologies (OpenStack, Proxmox, VMware).

• Experience with bare-metal provisioning and MAAS (Metal as a Service).

• Strong understanding of distributed systems and container orchestration.

• A process-oriented mindset with the ability to create SOPs and operational procedures from the ground up.

• Experience in incident response, escalation protocols, and on-call rotations.

• Capability to work independently in a fast-paced, engineering-focused environment.

• Strong technical skills aligned with team values.


🏝️ Benefits

• 100% remote work with flexible hours.

• High-impact role offering autonomy and ownership.

• Collaborative and international engineering team.

• Access to a cutting-edge tech stack with a strong emphasis on reliability and automation.

People also viewed

The CodestJul 26

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
IRIUMJul 26

Ingeniero/a Cloud DevOps

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€33k – €40k/year
ApplyView job
SólidesJul 26

Senior DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ResilincJul 25

Junior/Senior Site Reliability Engineer – Night Shift

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity GroupJul 25

Senior SRE / DevOps Engineer

Anywhere in the WorldFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
HOESSLER & HOESSLERJul 25

DevOps Software Engineer – Career Ambitions

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€65k – €75k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers