Staff DevOps Engineer

Posted Sep 16

This is a fully remote position, open to applicants in Colombia.

📋 Description

• Establish repeatable practices for service delivery through modular, reusable automation and a developer platform that facilitates self-service service delivery.

• Engage in governance controls aimed at minimizing risk and enhancing organizational standardization.

• Enhance enablement practices, which include design reviews, coordination of service launches, assessments of production readiness, definition and review of service-level objectives, incident management, and cost awareness.

• Assist delivery teams throughout their service lifecycle and levels of maturity.

• Collaborate with delivery teams to identify product-specific metrics and necessary remediations.

• Conduct system analysis, testing, and troubleshooting for faults.

• Address and prioritize escalations of production issues during off-hours as part of an on-call rotation.


⛳️ Requirements

• Over 8 years of engineering experience managing high-availability systems and supporting infrastructure in customer-facing production settings.

• Expertise in a high-level programming language, such as Python or Go.

• Proficient in Bash scripting and Linux environments.

• Familiarity with modern technical operating practices.

• Experience in system architecture and design.

• Proficient in continuous integration and continuous delivery methodologies using Jenkins, FluxCD, and GitHub Actions.

• Experience with Infrastructure as Code (IaC) utilizing Terraform, including Terraform modules.

• Knowledge of AWS cloud services and Kubernetes.

• Understanding of Site Reliability Engineering (SRE) principles and practices.

• Familiarity with Kubernetes cluster concepts and their design.

• Experience enhancing service observability through monitoring agents, metrics, logging, and dashboards.

• Knowledge of OpenTelemetry and Prometheus.

• Familiar with observability platforms such as Datadog, Splunk, Dynatrace, or Observe.

• Willingness to participate in an on-call rotation to handle production issue escalations during off-hours.

• Experience with AWS services and capabilities, including ECS, EKS, ECR, EC2, S3, RDS, VPCs, IAM policy documents, policies, roles, instance profiles, and CloudWatch Logs.

• Experience with Docker containers and container orchestration via ECS and EKS.

• Successful completion of employment verification and background checks.

• Verification of job titles and employment dates with two previous employers.


🏝️ Benefits

• Competitive salary paid in USD.

• 20 days of paid time off (PTO) annually.

• Opportunities for professional growth and career advancement.

• Company-supplied equipment.

• A collaborative, inclusive, and multicultural work environment.

• The chance to contribute to impactful projects alongside a skilled and supportive team.

People also viewed

FourEnergy GmbH11 hours ago

Senior DevOps Engineer – Operations

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ICF14 hours ago

Lead DevOps Engineer

US flagVirginia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$131.3k – $223.1k/year
ApplyView job
Mastercam18 hours ago

DevSecOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
C&S Informática22 hours ago

DevOps Engineer – Freelance/Contract, Mid-Level/Senior

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Convene1 day ago

Support and Deployment Engineer

SA flagSaudi Arabia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity Group1 day ago

SRE Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers