Staff DevOps Engineer

Posted Sep 8

This is a fully remote position, open to applicants in Serbia.

📋 Description

• Establish and guide the technical vision and quarterly roadmap for infrastructure, ensuring clear trade-offs and measurable objectives.

• Manage and enhance AWS and Kubernetes (EKS) infrastructure, focusing on cluster management, Karpenter autoscaling, Kyverno policy enforcement, and ensuring zero-downtime operations.

• Oversee Infrastructure as Code from start to finish using Terraform and AWS CDK in TypeScript.

• Take charge of GitLab CI/CD, including reusable/shared templates, OIDC, and self-managed GitLab.

• Develop and manage practical observability solutions using Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log pipelines, and APM.

• Ensure rapid blue-green deployments with health-gated automated rollback processes.

• Oversee zero-downtime PostgreSQL schema migrations utilizing expand/contract and CI migration gating.

• Lead security engineering initiatives in a HIPAA-compliant environment, encompassing secrets hygiene, credential rotation, leak scanning, PHI-aware log and data management, and Vault managed as code.

• Collaborate directly with product teams to eliminate infrastructure friction and enhance the developer experience.

• Leverage agentic AI as a fundamental component of workflows, integrating autonomous-agent outputs into production.

• Define technical direction, actively contribute to infrastructure deployment, and take ownership of reliability, cost, security, performance, deployment health, and developer velocity outcomes.


⛳️ Requirements

• Minimum of 6 years in DevOps/infrastructure engineering.

• Strong foundational knowledge of systems.

• Proficient in Linux administration and troubleshooting, including performance analysis, resource management, and process debugging.

• Hands-on experience with AWS services including EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, and S3.

• Practical experience with Kubernetes/EKS in production, covering cluster management, node scaling, and policy enforcement, including Karpenter, Kyverno, or similar tools.

• Strong skills in Infrastructure as Code using Terraform and AWS CDK in TypeScript.

• Ownership of CI/CD processes with GitLab CI/CD, reusable/shared templates, OIDC id_tokens, and self-managed GitLab.

• Experience in monitoring and observability using Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log shipping, and error tracking/APM.

• Background in security engineering with experience in secrets rotation, short-lived credentials, leak scanning, and PHI-aware logging.

• Familiarity with HashiCorp Vault as code, including KV, JWT/OIDC authentication for CI, and policy design.

• Experience with blue-green deployments featuring automated, health-gated rollback.

• Knowledge of PostgreSQL zero-downtime schema migrations using expand/contract and migration gating in CI.

• Experience with containers, including Docker, ECR, immutable tags, and image lifecycle management.

• Understanding of network and protocol fundamentals, including load balancing, TLS, and DNS.

• Hands-on experience with agentic AI workflows, such as Claude Code or similar technologies.

• Developer-focused mindset with strong problem-solving abilities for complex system challenges.

• Proficient in technical writing, including docs-as-code, ADRs, and design documents via MRs.

• Fluent in Russian and English (B2 level).

• Experience working effectively within remote, distributed teams.

• Preferred: experience in regulated or compliance-heavy environments, such as HIPAA or SOC 2.

• Preferred: familiarity with Ansible for VM fleet management.

• Preferred: experience with Node.js application operations using pm2 and npm.

• Preferred: knowledge of GitOps tooling such as ArgoCD or Flux.

• Preferred: deeper PostgreSQL database administration expertise.

• Preferred: AWS certifications are a plus.


🏝️ Benefits

• Competitive compensation package.

• Health insurance benefits following the probation period.

• Sports and wellness compensation.

• Personalized English lessons through Preply.

• 19 paid vacation days each year.

• 4 additional wellness days annually.

• Paid sick leave for the first 5 working days.

• Thoughtful gifts for significant life events.

• Offline corporate events.

• Fully remote long-term collaboration under a B2B model.

People also viewed

FourEnergy GmbH19 hours ago

Senior DevOps Engineer – Operations

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ICF21 hours ago

Lead DevOps Engineer

US flagVirginia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$131.3k – $223.1k/year
ApplyView job
Mastercam1 day ago

DevSecOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
C&S Informática1 day ago

DevOps Engineer – Freelance/Contract, Mid-Level/Senior

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Convene1 day ago

Support and Deployment Engineer

SA flagSaudi Arabia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity Group1 day ago

SRE Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers