Staff DevOps Engineer

Posted Sep 8

This is a fully remote position, open to applicants in Spain.

📋 Description

• Establish and lead the technical vision along with a quarterly roadmap for infrastructure, ensuring clear trade-offs and measurable objectives.

• Manage and enhance AWS and Kubernetes/EKS infrastructure, focusing on cluster management, Karpenter autoscaling, Kyverno policy enforcement, and ensuring zero-downtime operations.

• Take full ownership of Infrastructure as Code utilizing Terraform and AWS CDK in TypeScript.

• Oversee GitLab CI/CD processes, including the development of reusable/shared templates, OIDC, and self-managed GitLab.

• Develop and maintain effective observability solutions with Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log pipelines, and APM.

• Ensure swift blue-green deployments and health-gated automated rollback procedures.

• Manage zero-downtime PostgreSQL schema migrations along with CI migration gating.

• Lead security engineering efforts within a HIPAA-compliant environment, focusing on secrets hygiene, credential rotation, leak scanning, and PHI-aware log and data management, all managed as code with Vault.

• Collaborate with product engineering teams to eliminate infrastructure friction and enhance the developer experience.

• Integrate agentic AI as a primary workflow, incorporating autonomous-agent outputs into production systems.

• Define technical strategy, directly engage in infrastructure deployment, and take ownership of reliability, cost, security, performance, deployment health, and developer velocity outcomes.


⛳️ Requirements

• Minimum of 6 years of experience in DevOps/infrastructure engineering.

• Strong foundational knowledge of systems and proficient in Linux administration and troubleshooting, including performance analysis, resource management, and process debugging.

• Practical experience with AWS services such as EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, and S3.

• Hands-on experience managing production Kubernetes/EKS environments, including cluster management, node scaling, and policy enforcement, with tools like Karpenter and Kyverno.

• Extensive experience in Infrastructure as Code using Terraform and AWS CDK in TypeScript.

• Proven ownership of CI/CD processes using GitLab CI/CD, including reusable/shared templates, OIDC id_tokens, and self-managed GitLab.

• Experience in practical monitoring and observability with tools such as Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log-shipping, and error tracking/APM.

• Solid experience in security engineering, covering secrets rotation, short-lived credentials, leak scanning, and PHI-aware logging.

• Familiarity with HashiCorp Vault as code, including KV, JWT/OIDC authentication for CI, and policy design.

• Experience executing blue-green deployments and implementing automated, health-gated rollback strategies.

• Knowledge of PostgreSQL zero-downtime schema migrations utilizing expand/contract and CI migration gating.

• Proficiency with Docker, ECR, immutable tags, and managing image lifecycles.

• Understanding of network/protocol fundamentals, including load balancing, TLS, and DNS.

• Hands-on experience with agentic AI workflows, such as Claude Code or similar technologies.

• A developer-oriented mindset with strong problem-solving skills for intricate system challenges, along with excellent technical writing capabilities, including docs-as-code, ADRs, and design documents via MRs.

• Proficiency in both Russian and English (B2 level).

• Experience effectively collaborating in remote, distributed teams.

• Experience in a regulated or compliance-heavy environment, such as HIPAA or SOC 2, is advantageous.

• Familiarity with configuration management using Ansible is a plus.

• Experience with Node.js application operations using pm2 and npm is an advantage.

• Knowledge of GitOps tools such as ArgoCD or Flux and advanced PostgreSQL database administration are beneficial.

• AWS certifications will be considered a plus.


🏝️ Benefits

• Competitive salary package.

• Health insurance coverage after the probation period.

• Compensation for sports and wellness activities.

• Personalized English language lessons through Preply.

• 19 paid vacation days per year.

• 4 additional wellness days annually.

• Paid sick leave for the initial 5 working days.

• Thoughtful gifts for significant life events.

• Offline corporate events.

• Fully remote long-term collaboration under a B2B model.

People also viewed

FourEnergy GmbH13 hours ago

Senior DevOps Engineer – Operations

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ICF15 hours ago

Lead DevOps Engineer

US flagVirginia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$131.3k – $223.1k/year
ApplyView job
Mastercam19 hours ago

DevSecOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
C&S Informática1 day ago

DevOps Engineer – Freelance/Contract, Mid-Level/Senior

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Convene1 day ago

Support and Deployment Engineer

SA flagSaudi Arabia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity Group1 day ago

SRE Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers