Remotery

Staff DevOps Engineer

Posted Aug 5

This is a fully remote position, open to applicants in Europe.

📋 Description

• Establish and drive the technical vision and quarterly roadmap for infrastructure, encompassing clear trade-offs and measurable objectives.

• Manage and enhance AWS and Kubernetes (EKS) infrastructure, which includes cluster management, autoscaling, policy enforcement, and ensuring zero-downtime operations.

• Take ownership of Infrastructure as Code utilizing Terraform and AWS CDK in TypeScript.

• Oversee GitLab CI/CD, incorporating reusable/shared templates, OIDC, and self-managed GitLab.

• Develop and maintain practical observability using Prometheus, Grafana, OpenTelemetry, OpenSearch, and CloudWatch.

• Sustain log pipelines and APM to provide a unified view of system health.

• Execute blue-green deployments with health-gated automated rollback procedures.

• Manage zero-downtime PostgreSQL schema migrations and CI migration gating.

• Lead security engineering within a HIPAA-compliant environment, focusing on secrets hygiene, PHI-aware logging and data management, leak scanning, and Vault managed as code.

• Collaborate directly with product teams to alleviate infrastructure friction and enhance developer experience.

• Leverage agentic AI as a foundational workflow, integrating autonomous-agent output into production processes.

• Define technical direction, implement infrastructure changes hands-on, and manage outcomes related to reliability, cost, security, performance, deployment health, and developer velocity.


⛳️ Requirements

• Over 6 years of experience in DevOps/infrastructure engineering.

• Strong understanding of systems fundamentals, Linux administration, and troubleshooting, including performance analysis, resource management, and process debugging.

• Practical AWS experience with EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, and S3.

• Experience with production Kubernetes/EKS, including cluster management, node scaling, and policy enforcement; familiarity with Karpenter, Kyverno, or similar tools.

• Strong expertise in Infrastructure as Code with Terraform and AWS CDK in TypeScript.

• Ownership of CI/CD processes using GitLab CI/CD, including reusable/shared templates, OIDC id_tokens, and self-managed GitLab.

• Hands-on monitoring and observability experience with Prometheus, Grafana, OpenTelemetry, OpenSearch, and CloudWatch; including log shipping and error tracking/APM.

• Practical experience in security engineering, focusing on secrets rotation, short-lived credentials, leak scanning, and PHI-aware logging.

• Experience with HashiCorp Vault as code, including KV, JWT/OIDC authentication for CI, and policy design.

• Familiarity with blue-green deployments and automated, health-gated rollback.

• Experience with PostgreSQL zero-downtime schema migration using expand/contract and CI migration gating.

• Knowledge of containerization with Docker, ECR, immutable tags, and image lifecycle management.

• Understanding of network and protocol fundamentals, including load balancing, TLS, and DNS.

• Hands-on experience with agentic AI workflows, such as Claude Code or similar, including delegating to autonomous agents and integrating their output into production.

• A developer-focused mindset with strong problem-solving abilities for complex system challenges.

• Proficient technical writing skills, including docs-as-code, ADRs, and design documentation via MRs.

• Fluent in Russian and English (B2 level).

• Ability to work effectively in remote, distributed teams.

• Preferred: experience in HIPAA, SOC 2, or similar regulated environments; familiarity with Ansible; Node.js operations; GitOps tooling; advanced PostgreSQL administration; AWS certifications.


🏝️ Benefits

• Competitive compensation package.

• Fully remote long-term collaboration under a B2B model.

• Health insurance provided after the probation period.

• Sports & wellness compensation available.

• Personalized English lessons through Preply.

• 19 paid vacation days each year.

• 4 additional wellness days annually.

• Paid sick leave for the first 5 working days.

• Thoughtful gifts for significant life events.

• Offline corporate events.

People also viewed

CWILL17 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3718 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT18 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group18 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo19 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch19 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers