
Kubernetes Engineer
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in New York.
• Analyze production-focused Kubernetes scenarios in both managed and self-hosted environments.
• Evaluate cluster configurations, workload behaviors, and operational readiness.
• Assess Kubernetes networking, service discovery, CNI, DNS, ingress, and service connectivity.
• Examine PersistentVolume, PersistentVolumeClaim, storage classes, volume lifecycles, provisioning, mounting, and persistence challenges.
• Review RBAC configurations, roles, cluster roles, bindings, service-account permissions, and decisions surrounding least-privilege access.
• Diagnose issues such as CrashLoopBackOff, OOMKilled, scheduling failures, pod eviction, connectivity problems, routing, and other Kubernetes-related failures.
• Create and review Kubernetes manifests while assessing Helm charts for accuracy, maintainability, and deployment safety.
• Analyze resource specifications, probes, requests, limits, selectors, dependencies, and templating.
• Evaluate live-cluster troubleshooting scenarios, diagnostic reasoning, incident-response strategies, prioritization, escalation, and remediation efforts.
• Assess considerations related to scaling, availability, resilience, observability, autoscaling, service mesh, and production readiness.
• Review Prometheus and Grafana for monitoring, alerting, metrics, logs, operational signals, HPA configurations, and Istio scenarios.
• Evaluate Kubernetes tasks against structured technical criteria and deliver clear, rubric-based written feedback.
• Reference configuration, runtime behavior, and diagnostic evidence to differentiate valid alternatives from incorrect solutions.
• Minimum of 3 years of hands-on experience with production Kubernetes environments.
• Experience managing EKS, GKE, AKS, or self-managed Kubernetes clusters.
• In-depth knowledge of CNI, DNS, ingress, PV/PVC storage, RBAC, scheduling, and cluster failure modes.
• Extensive experience in authoring and reviewing Kubernetes manifests and Helm charts.
• Proven track record in debugging live incidents in production clusters.
• Proficiency in programming languages such as Go, Python, or TypeScript.
• Strong understanding of workload resource management and ensuring production reliability.
• CKA or CKAD certification is preferred.
• Familiarity with Istio, HPA, Prometheus, Grafana, or similar technologies is a plus.
• Previous experience in SRE, platform engineering, code review, or technical task grading is advantageous.
• Excellent written communication skills with the ability to provide precise technical feedback.
• Work must be conducted without utilizing confidential or proprietary information from any employer, client, institution, or other third party.
• H1-B and STEM OPT support is not available for this position.
• Part-time independent contractor position.
• Fully remote work within the United States.
• Flexible scheduling based on project requirements.
• Compensation ranges from $60 to $80 per hour.
• Project durations may be extended, shortened, or concluded based on needs and performance.
RockstarDevelopers GmbH
DeweyLearn, Inc.
InvestEngine
Get handpicked remote jobs straight to your inbox weekly.