AI Infrastructure & Platform Operations Engineer

atMirantisRemoteEuropeFull-timeInfrastructure EngineerMid-levelSenior$60k – $67k/year

Posted Jul 7

This is a fully remote position, open to applicants in Europe.

πŸ“‹ Description

β€’ Oversee, manage, and provide support for production AI infrastructure platforms.

β€’ Analyze and resolve incidents related to infrastructure, networking, hardware, and platforms.

β€’ Provide support for NVIDIA GPU infrastructure and related platform services.

β€’ Monitor and troubleshoot environments based on Kubernetes.

β€’ Investigate issues concerning performance, availability, and reliability across infrastructure and platform components.

β€’ Collaborate with engineering teams, hardware vendors, datacenter staff, and service delivery teams to address technical challenges.

β€’ Engage in incident response, root cause analysis, and initiatives for operational improvements.

β€’ Contribute to enhancements in monitoring, observability, automation, and operational processes.

β€’ Maintain operational documentation, runbooks, and knowledge articles.


⛳️ Requirements

β€’ A minimum of 3 years of experience in infrastructure operations, platform operations, network operations, site reliability engineering, cloud operations, datacenter operations, or similar technical positions.

β€’ Proficient in Linux administration and troubleshooting.

β€’ Solid understanding of networking concepts with experience in diagnosing infrastructure-related issues.

β€’ Familiarity with Kubernetes in production settings.

β€’ Experience in supporting production infrastructure and services.

β€’ Strong analytical and problem-solving capabilities.

β€’ Experience adhering to structured operational and incident management processes.

β€’ Exceptional communication and collaboration abilities.

β€’ Capability to operate within a shift-based operational framework.

β€’ Experience in one or more of the following areas is highly desirable: NVIDIA GPU infrastructure and accelerated computing platforms, InfiniBand networking and NVIDIA UFM, Kubernetes platform operations, AI infrastructure or HPC environments, Site Reliability Engineering (SRE) or Platform Engineering, Observability platforms such as Grafana, Prometheus, ELK, or OpenTelemetry, Infrastructure automation technologies and Infrastructure-as-Code practices, Large-scale distributed systems and production platforms.


🏝️ Benefits

β€’ Work with some of the most advanced AI infrastructure environments currently in production.

β€’ Gain exposure to NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments.

β€’ Contribute to defining the operational and support frameworks for next-generation AI infrastructure.

β€’ Be part of a team that is shaping the future of AI-powered operations through k0rdent AI.

β€’ Join a growing organization that is heavily investing in AI infrastructure and platform services.

People also viewed

PSI CRO AG6 hours ago

IT Infrastructure Engineer – Windows, Active Directory, VMware

EE flagEstonia OnlyFull-timeInfrastructure Engineer
ApplyView job
PSI CRO AG6 hours ago

IT Infrastructure Engineer, Windows, Active Directory, VMware

CZ flagCzechia OnlyFull-timeInfrastructure Engineer
ApplyView job
PSI CRO AG6 hours ago

IT Infrastructure Engineer – Windows, Active Directory, VMware

LV flagLatvia OnlyFull-timeInfrastructure Engineer
ApplyView job
PSI CRO AG7 hours ago

IT Infrastructure Engineer – Windows, Active Directory, VMware

PT flagPortugal OnlyFull-timeInfrastructure Engineer
ApplyView job
Delinea12 hours ago

Senior Cloud Engineer – Infrastructure Services, FedRAMP

US flagUnited States OnlyFull-timeInfrastructure Engineer$130k – $160k/year
ApplyView job
Coinbase12 hours ago

Senior Software Engineer, Developer Infrastructure

US flagUnited States OnlyFull-timeInfrastructure Engineer$186.1k – $218.9k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers