Remotery

Senior DevOps Engineer

Posted Jul 21

This is a fully remote position, open to applicants in Cyprus.

📋 Description

• Take full ownership of one of our essential platform areas, either observability (VictoriaMetrics, Grafana, Graylog / VictoriaLogs, fluent bit, exporters, alerting) or CI/CD (Jenkins scripted pipelines, Harbor, Nexus, build agents) — you will lead its architecture, reliability, and strategic roadmap.

• Spearhead technical projects from start to finish: gather requirements, draft the design document, break down tasks, implement solutions, deliver to production, and maintain operational health thereafter.

• Bring clarity to uncertain situations by clearly defining requirements, assumptions, and next steps.

• Design systems for reliability and scalability: enhance the architecture of our platforms — including topology, integration points, scaling strategies, and reliability models.

• Provide support to developers: deploy and monitor applications on both on-premise servers and Kubernetes (Helm), troubleshoot builds and deployments, assist teams with metrics, alerts, and logs; participate in chat duty in developer support channels.

• Eliminate manual tasks through automation: repetitive operations, provisioning, and maintenance should be codified rather than executed manually.

• Investigate production incidents as the senior escalation point for your domain: lead resolution efforts, conduct post-mortems, and implement systemic improvements. Participate in on-call rotations and enhance the operational standards for on-call responsibilities.

• Mentor junior engineers through design discussions, code reviews, and collaborative work; identify and mitigate debt-inducing shortcuts during the review process.

• Leverage AI in all facets of daily tasks: engaging in research, troubleshooting, and development.


⛳️ Requirements

• Minimum of 6 years of experience as a DevOps Engineer / SRE (or closely related responsibilities).

• Proven track record of managing technical initiatives from inception to completion — encompassing requirements gathering, technical design, and production delivery. You should be able to highlight initiatives that you owned, rather than merely tasks you completed.

• Proficient Linux skills (we utilize Ubuntu).

• Familiarity with the Prometheus stack: understanding of metric types, exporters, and alerting mechanisms — sufficient to navigate and enhance an existing setup.

• Practical experience with CI/CD: designing pipelines, orchestrating builds, and delivering artifacts.

• Proficiency in containerization: Docker, image creation, and registries.

• Experience with Ansible.

• Proficient in Git.

• Skills in Bash or Python scripting for automation and observability (such as writing exporters and reducing routine tasks).

• Experience in production/on-call roles: diagnosing incidents, restoring services, and leading post-mortems.

• Background in mentoring less experienced engineers.

• Strong sense of ownership and meticulous attention to detail. Downtime is costly: during peak events, just 10 minutes of downtime can lead to losses of around €500k.


🏝️ Benefits

• 31 days of vacation

• Fully covered telemedicine plan

• Home Office Setup Assistance: the company provides support for purchasing furniture (office chair, office desk, monitor) and other items to establish a comfortable workspace

• English language courses

• Relevant professional development opportunities

• Access to gym or swimming pool

• Co-working options

• Flexibility for remote work

People also viewed

CWILL21 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3722 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT22 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group22 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo23 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch23 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers