
Service Delivery Manager
Posted Aug 21

Posted Aug 21
This is a fully remote position, open to applicants in Kazakhstan.
β’ Lead the engineering team responsible for delivering Mirantis's Neocloud Infrastructure-as-a-Service.
β’ Oversee engineers tasked with large-scale infrastructure and platform operations.
β’ Manage rostering and coverage for a globally distributed 24x7 operational model.
β’ Conduct one-on-one meetings, set goals, promote skills development, and carry out performance reviews.
β’ Identify and address team skill deficiencies as the scope of work increases.
β’ Facilitate onboarding of engineers into active customer environments.
β’ Take ownership of service delivery performance in relation to contracted SLAs.
β’ Define, monitor, and report service KPIs, including availability, MTTR, MTTA, ticket aging/backlog, and customer satisfaction.
β’ Chair both internal and customer-facing service review meetings.
β’ Maintain and enhance runbooks, escalation paths, and operational documentation.
β’ Manage SLA risks and escalate the commercial impact to leadership.
β’ Oversee the incident lifecycle from detection to post-incident review.
β’ Serve as the primary escalation point for major and critical incidents/outages.
β’ Coordinate cross-functional resources and communicate incident status updates to customers and leadership.
β’ Conduct blameless post-incident reviews and ensure tracking of corrective actions to closure.
β’ Analyze incident trends to promote proactive reliability enhancements.
β’ Ensure that on-call and escalation rotas are adequately staffed, documented, and tested.
β’ Act as the senior operational contact for key customers.
β’ Collaborate with Sales and Solutions Architecture during the customer onboarding process.
β’ Represent service delivery performance and improvement strategies in customer business reviews.
β’ Enhance monitoring, observability, and incident management tools.
β’ Maintain onboarding and handover documentation.
β’ Ensure operational readiness reviews are conducted prior to new customer environments going live.
β’ Demonstrated experience managing technical operations or managed services teams responsible for large-scale infrastructure and/or platform operations.
β’ A technical background sufficient to understand incident complexities and lead responses with credibility.
β’ Direct experience in owning incident management and major incident/outage responses, including customer communication during live incidents.
β’ Strong understanding of SLA frameworks, service reporting, and escalation management.
β’ Experience managing distributed or shift-based teams operating on a 24x7 rotation.
β’ Excellent stakeholder management skills, with the ability to communicate effectively with engineers and customer executives.
β’ A data-driven approach to service management, including the ability to build and present operational metrics.
β’ Must be based in the EU and eligible to work there.
β’ Experience with observability tools and practices (monitoring, logging, tracing, alerting) is a significant advantage.
β’ Familiarity with infrastructure and platform technologies.
β’ ITIL or equivalent service management certification.
β’ Experience in a neocloud, hyperscaler, colocation, or managed hosting environment.
β’ Experience managing services for customers who possess their own physical infrastructure.
β’ Experience leading teams whose responsibilities expanded from infrastructure into platform-layer operations.
β’ Opportunities for professional development and training.
β’ Attend conferences and participate in working groups.
β’ Company outings, happy hours, hackathons, and tech talks.
β’ Competitive compensation package complemented by a robust benefits plan.
β’ Flexible remote work arrangement (employees can work remotely).
Collibra
WNS
OLLY PBC
Woodard & Curran
Get handpicked remote jobs straight to your inbox weekly.