
Senior DevOps Engineer
Posted Jul 21

Posted Jul 21
This is a fully remote position, open to applicants in Cyprus.
• Take full ownership of one of our essential platform areas, either observability (VictoriaMetrics, Grafana, Graylog / VictoriaLogs, fluent bit, exporters, alerting) or CI/CD (Jenkins scripted pipelines, Harbor, Nexus, build agents) — you will lead its architecture, reliability, and strategic roadmap.
• Spearhead technical projects from start to finish: gather requirements, draft the design document, break down tasks, implement solutions, deliver to production, and maintain operational health thereafter.
• Bring clarity to uncertain situations by clearly defining requirements, assumptions, and next steps.
• Design systems for reliability and scalability: enhance the architecture of our platforms — including topology, integration points, scaling strategies, and reliability models.
• Provide support to developers: deploy and monitor applications on both on-premise servers and Kubernetes (Helm), troubleshoot builds and deployments, assist teams with metrics, alerts, and logs; participate in chat duty in developer support channels.
• Eliminate manual tasks through automation: repetitive operations, provisioning, and maintenance should be codified rather than executed manually.
• Investigate production incidents as the senior escalation point for your domain: lead resolution efforts, conduct post-mortems, and implement systemic improvements. Participate in on-call rotations and enhance the operational standards for on-call responsibilities.
• Mentor junior engineers through design discussions, code reviews, and collaborative work; identify and mitigate debt-inducing shortcuts during the review process.
• Leverage AI in all facets of daily tasks: engaging in research, troubleshooting, and development.
• Minimum of 6 years of experience as a DevOps Engineer / SRE (or closely related responsibilities).
• Proven track record of managing technical initiatives from inception to completion — encompassing requirements gathering, technical design, and production delivery. You should be able to highlight initiatives that you owned, rather than merely tasks you completed.
• Proficient Linux skills (we utilize Ubuntu).
• Familiarity with the Prometheus stack: understanding of metric types, exporters, and alerting mechanisms — sufficient to navigate and enhance an existing setup.
• Practical experience with CI/CD: designing pipelines, orchestrating builds, and delivering artifacts.
• Proficiency in containerization: Docker, image creation, and registries.
• Experience with Ansible.
• Proficient in Git.
• Skills in Bash or Python scripting for automation and observability (such as writing exporters and reducing routine tasks).
• Experience in production/on-call roles: diagnosing incidents, restoring services, and leading post-mortems.
• Background in mentoring less experienced engineers.
• Strong sense of ownership and meticulous attention to detail. Downtime is costly: during peak events, just 10 minutes of downtime can lead to losses of around €500k.
• 31 days of vacation
• Fully covered telemedicine plan
• Home Office Setup Assistance: the company provides support for purchasing furniture (office chair, office desk, monitor) and other items to establish a comfortable workspace
• English language courses
• Relevant professional development opportunities
• Access to gym or swimming pool
• Co-working options
• Flexibility for remote work
CWILL
a37
GT
Sigma Software Group
Get handpicked remote jobs straight to your inbox weekly.