
Senior Kubernetes Engineer
Posted Sep 18

Posted Sep 18
This is a fully remote position, open to applicants in Egypt.
• Design and manage production Kubernetes clusters on bare metal, private cloud, and on-premises infrastructure.
• Take ownership of the Kubernetes platform architecture from cluster design to deployment and ongoing operations.
• Define reference Kubernetes deployments that can be installed and operated by client infrastructure teams.
• Design and oversee CNI networking, CSI integrations, persistent storage, ingress, and load balancing.
• Manage etcd backup, recovery, health monitoring, and disaster recovery processes.
• Develop and maintain deployment workflows utilizing Helm and GitOps tools such as ArgoCD or Flux.
• Employ Terraform to automate infrastructure and Kubernetes platform provisioning when applicable.
• Design and implement air-gapped and restricted-network Kubernetes deployments.
• Oversee container image mirroring, private registries, and offline installation methods.
• Deploy and manage stateful workloads including Kafka, Flink, Spark, databases, and their operators.
• Create highly available Kubernetes platforms for mission-critical workloads.
• Implement security hardening for regulated and security-sensitive environments.
• Define and enforce RBAC, network policies, secrets management, and cluster security policies.
• Plan for cluster capacity, node resources, scaling strategies, and workload placement.
• Design and execute disaster recovery and business continuity plans.
• Plan and execute Kubernetes upgrades with minimal or no downtime.
• Support GPU scheduling and node management for AI and data-intensive workloads.
• Collaborate with client infrastructure and operations teams during deployment, handover, and operational readiness.
• Engage in security reviews and technical architecture discussions.
• Diagnose complex networking, storage, scheduling, and cluster-level production problems.
• Establish operational standards, runbooks, monitoring requirements, and troubleshooting protocols.
• Demonstrated experience as a Senior Kubernetes Engineer, Platform Engineer, DevOps Engineer, or a related role.
• Strong practical experience in building and managing production Kubernetes clusters on bare metal or private cloud.
• Familiarity with client-owned, on-premises, or sovereign infrastructure environments.
• Solid understanding of Kubernetes architecture and the components typically abstracted away by managed Kubernetes services.
• Extensive knowledge of Kubernetes networking and CNI behavior.
• Significant experience with CSI, persistent storage, and stateful workloads on Kubernetes.
• Strong comprehension of ingress controllers, service networking, and load balancing without reliance on a public cloud provider.
• Practical experience with etcd operations, including backup, recovery, troubleshooting, and cluster health management.
• Experience with Helm and Kubernetes package management.
• Practical exposure to ArgoCD, Flux, or similar GitOps tools.
• Strong Terraform skills for infrastructure and platform provisioning.
• Proven experience deploying Kubernetes in air-gapped or highly restricted network scenarios.
• Familiarity with private container registries, image mirroring, and offline installation workflows.
• Experience operating stateful technologies on Kubernetes, including Kafka, Flink, Spark, databases, and their operators.
• Strong understanding of Kubernetes security, including RBAC, network policies, secrets management, and policy enforcement.
• Experience in hardening Kubernetes platforms for regulated or security-sensitive environments.
• Strong understanding of PKI, certificates, TLS, and certificate lifecycle management within Kubernetes environments.
• Experience in capacity planning, high availability, disaster recovery, and cluster scaling.
• Proven track record of performing Kubernetes upgrades while minimizing service disruptions.
• Experience with GPU scheduling, node management, and Kubernetes workloads that support AI or data-intensive applications.
• Strong troubleshooting abilities across networking, storage, compute, scheduling, and Kubernetes control-plane components.
• Experience working directly in client data centers and participating in infrastructure or security assessments.
• Ability to create clear deployment documentation, operational runbooks, and handover materials for client operations teams.
• Strong understanding of production reliability, observability, and operational readiness for Kubernetes platforms.
• Competitive salary and benefits package.
• Opportunities for professional growth and development.
• Flexible working hours and remote work options.
• Collaborative and innovative working environment.
Oscilar
Veradigm®
Get handpicked remote jobs straight to your inbox weekly.