
Senior Platform Engineer, Network Infrastructure
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in India.
• Design, develop, and manage the Kubernetes platform that supports GNI's network automation, telemetry, and operations across data centers, colocation, and cloud environments.
• Take ownership of lifecycle management for GNI Kubernetes environments, which includes cluster onboarding, upgrades, capacity management, availability, and recovery.
• Create production-grade software and automation tools for cluster provisioning, validation, upgrades, remediation, and secure multi-cluster delivery utilizing GitOps.
• Offer production support for network services that are hosted on the platform, collaborating with Network Automation and service teams who maintain ownership of application architecture, code, and features.
• Troubleshoot complex failures within the Kubernetes platform and hosted services, focusing on control-plane health, cluster networking, storage, scheduling, workload placement, and multi-cluster dependencies.
• Lead issues from the initial signal to confirmed resolution.
• Establish production-readiness and observability standards for the platform and hosted network services, which encompass health signals, capacity, alerts, runbooks, and recovery processes.
• Engage in CFR’s production on-call rotation, which includes scheduled after-hours and weekend duties.
• Oversee incident response and recovery efforts, ensuring completion of corrective actions.
• Bachelor’s degree in Computer Science, Engineering, or a related discipline, or equivalent professional experience.
• Over 8 years of experience in building or managing production Kubernetes platforms, network infrastructure, or distributed systems.
• Extensive experience with Kubernetes at scale, covering cluster lifecycle management, upgrades, networking, storage, and recovery.
• Proficient in at least one general-purpose programming language, such as Go or Python.
• Familiarity with GitOps, infrastructure as code, CI/CD pipelines, and automated production deployment.
• Background in deploying and supporting network automation or telemetry services on Kubernetes.
• Experience with production on-call duties, incident response, root-cause analysis, and ensuring completion of corrective actions.
• Health insurance
• Retirement plans
• Paid time off
• Flexible work arrangements
• Professional development
Blue Ocean Global Technology
ShiftKey
nbn® Australia
Mirantis
Get handpicked remote jobs straight to your inbox weekly.