Remotery

Senior DevOps Engineer – Kubernetes, GitOps

Posted Aug 5

This is a fully remote position, open to applicants in Romania.

📋 Description

• Design, manage, and enhance production-grade Kubernetes environments

• Develop and sustain Kubernetes workloads, manifests, Helm charts, Kustomize overlays, and deployment strategies

• Implement and facilitate GitOps workflows utilizing Flux, Argo CD, or comparable tools

• Enhance CI/CD pipelines for dependable, repeatable, and secure application delivery

• Diagnose production challenges across Kubernetes, networking, compute, storage, and application layers in collaboration with engineering teams

• Create automation solutions for infrastructure provisioning, deployment, monitoring, and incident management

• Support cloud infrastructure operations across AWS, GCP, Windows, or hybrid settings

• Oversee Kubernetes networking, ingress, service mesh, DNS, TLS, secrets, and workload identity frameworks

• Promote Infrastructure as Code standards using Terraform, Pulumi, CloudFormation, or similar platforms

• Partner with security teams to enhance container, cluster, secrets, and supply-chain security

• Guide engineers on Kubernetes, DevOps, GitOps, and best practices for production operations


⛳️ Requirements

• Extensive hands-on experience with Kubernetes in a production environment

• Advanced understanding of Kubernetes, including deployments, services, ingress, config maps, secrets, RBAC, network policies, probes, HPA, and storage solutions

• Proficient in GitOps methodologies using Flux or Argo CD

• Strong experience with CI/CD tools such as GitHub Actions, GitLab CI, Jenkins, CircleCI, or similar platforms

• Familiarity with Helm and/or Kustomize

• Solid skills in Linux systems, networking, and troubleshooting

• Experience with cloud infrastructure, ideally AWS, encompassing EKS, IAM, VPC, Route 53, ALB/NLB, S3, RDS, or CloudWatch

• Proficiency in Infrastructure as Code using Terraform or equivalent tools

• Strong scripting abilities in Bash, Python, Go, or related languages

• Experience with observability tools such as Prometheus, Grafana, OpenTelemetry, Datadog, New Relic, Better Stack, or similar

• Capability to troubleshoot intricate production issues across application, infrastructure, and networking layers

• Excellent written communication skills for creating runbooks, architecture documentation, and operational guides

• Experience with Service Mesh is a plus

• Familiarity with Secrets Management is a plus

• Background in Platform Engineering is a plus

• Experience with multi-region architectures is a plus

• Knowledge of High Availability and Disaster Recovery solutions is a plus

• Familiarity with Microservices and Distributed Systems is a plus

• Experience in Incident Management and Postmortem practices is a plus

• Knowledge of SRE practices and SLO-based operations is a plus


🏝️ Benefits

• Health insurance

• Relocation program

• Flexibility for remote work

• Opportunities for professional development

• Access to certification programs

• Investment in mentorship and talent programs

• Opportunities for internal mobility

• Internship opportunities available

• Work on significant projects for leading global clients

• A diverse and supportive multicultural work environment

• Regular team-building social events within the company

People also viewed

CWILL17 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3718 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT18 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group18 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo19 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch19 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers