
Senior DevOps Engineer
Posted Jul 21

Posted Jul 21
This is a fully remote position, open to applicants in Turkey.
• Take full ownership of one of our key platform areas: either observability (including VictoriaMetrics, Grafana, Graylog / VictoriaLogs, fluent bit, exporters, and alerting) or CI/CD (featuring Jenkins scripted pipelines, Harbor, Nexus, and build agents) — you will lead its architecture, ensure its reliability, and define its roadmap.
• Spearhead technical initiatives from start to finish: collect requirements, draft the design document, break down tasks, implement solutions, deliver to production, and maintain operational health thereafter.
• Bring clarity to uncertain situations by articulating requirements, assumptions, and subsequent steps.
• Design systems for reliability and scalability: enhance the architecture of our platforms, focusing on topology, integration points, scaling strategies, and reliability models.
• Provide support to developers: deploy and monitor applications on both on-premise servers and Kubernetes (using Helm), troubleshoot builds and deployments, assist teams with metrics, alerts, and logs; engage in chat duty within developer support channels.
• Automate repetitive tasks: ensure that operations, provisioning, and maintenance are codified rather than executed manually.
• Investigate production incidents as the senior escalation point in your area: drive resolutions, lead post-mortems, and implement systemic improvements. Participate in on-call rotations and elevate the standards for on-call practices.
• Mentor junior engineers through design discussions, code reviews, and collaborative work; identify and address debt-inducing shortcuts during the review process.
• Leverage AI in all aspects of daily operations: from research and troubleshooting to development.
• A minimum of 6 years of experience as a DevOps Engineer or SRE (or similar roles with closely related responsibilities).
• Proven experience managing technical initiatives from inception to completion — from gathering requirements and creating technical designs to delivering production-ready solutions. You should be able to highlight initiatives that you have owned, rather than merely tasks you executed.
• Strong proficiency in Linux (we utilize Ubuntu).
• Familiarity with the Prometheus stack: understanding metric types, exporters, and alerting mechanisms — enough expertise to navigate and extend an existing setup.
• Practical experience in CI/CD: including pipeline design, build orchestration, and artifact delivery.
• Knowledge of containerization technologies: Docker, image building, and registries.
• Proficiency with Ansible.
• Experience with Git.
• Competence in Bash or Python scripting aimed at automation and observability (such as writing exporters and minimizing routine tasks).
• Experience in production/on-call environments: diagnosing incidents, restoring services, and leading post-mortems.
• Experience mentoring less experienced engineers.
• Strong sense of ownership and attention to detail. Downtime is costly: during peak events, just 10 minutes of downtime could result in approximately $500k in losses.
• 31 days of paid time off.
• Fully covered telemedicine plan.
• Home Office Setup Assistance: the company provides support for purchasing furniture (like office chair, desk, monitor) and other items to create an optimal workspace.
• English language learning courses.
• Opportunities for relevant professional education.
• Access to gym or swimming pool facilities.
• Co-working space availability.
• Flexible remote working options.
DATAGROUP
Ambush
DuoKey
TEKsystems
Get handpicked remote jobs straight to your inbox weekly.