Senior DevOps Engineer

Posted 1 day ago

This is a fully remote position, open to applicants in United States, +1 more country.

📋 Description

• Design, construct, and manage AWS infrastructure, including production workloads on Amazon EKS and ECS/Fargate.

• Take ownership of infrastructure projects through various stages such as design, testing, deployment, rollback planning, documentation, and operational handoff.

• Create infrastructure as code, automate deployments, and develop reusable platform capabilities using Terraform, Ansible, Helm, and Argo CD.

• Establish and sustain deployment pipelines and automation, collaborating with Development on application modernization, scaling, validation, production readiness, and rollback processes.

• Diagnose infrastructure and application issues across operating systems, networking, containers, application runtimes, and database dependencies.

• Enhance monitoring, performance, and cost efficiency; implement security measures in collaboration with Information Security; oversee platform upgrades and patching; and test backup restoration and disaster recovery processes.

• Contribute to technical standards and design evaluations, mentor engineers, and coordinate with management to prioritize risks and improvements.

• Engage in the on-call rotation, lead incident diagnosis, and execute planned maintenance activities.

• Communicate effectively during incidents and utilize follow-up tasks, automation, and runbooks to minimize recurring issues.


⛳️ Requirements

• Extensive infrastructure and DevOps experience, generally eight or more years in infrastructure and five or more years in automation-focused roles, or equivalent demonstrated expertise.

• Strong hands-on experience with production Kubernetes, including upgrades, networking, access controls, scaling, and troubleshooting.

• Experience with Amazon EKS is highly preferred.

• Comprehensive AWS experience across networking, IAM, compute, storage, load balancing, and managed databases.

• Practical knowledge of Terraform, containers, Helm, and Git-based deployment workflows, including the construction and maintenance of CI/CD pipelines using GitHub Actions, GitLab CI/CD, or similar tools.

• Proficient in Linux administration and troubleshooting skills, encompassing DNS, HTTP/TLS, and system performance.

• Strong networking and troubleshooting capabilities, including routing, firewalls, and secure connectivity such as site-to-site VPNs or SSH tunnels.

• Capability to write maintainable automation in Python, Go, PowerShell, or another suitable language, in addition to shell scripting.

• Experience in diagnosing application and database issues using metrics, logs, traces, and monitoring tools.

• Proven ability to independently deliver complex tasks, implement safe production changes, and communicate technical decisions effectively across teams.

• A degree in Computer Science or a related field, or equivalent practical experience.


🏝️ Benefits

• Competitive salary and performance-based incentives.

• Comprehensive health, dental, and vision insurance.

• Generous paid time off and holiday schedule.

• Opportunities for professional development and continuing education.

• Flexible working hours and remote work options.

People also viewed

Koniag Government Services23 hours ago

Architect/DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
FP Markets (First Prudential Markets)1 day ago

Senior DevOps Engineer

AM flagArmenia, +4 more countriesFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
InRule1 day ago

Site Reliability Engineer

US flagUnited States, +1 more countryFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Thumbtack1 day ago

Senior Software Engineer, Site Reliability Engineering

US flagUnited States, +38 more locationsFull-timeDevOps & Site Reliability Engineer (SRE)$179.4k – $272.8k/year
ApplyView job
Thumbtack1 day ago

Senior Software Engineer, Site Reliability Engineering

CA flagCanada, +1 more countryFull-timeDevOps & Site Reliability Engineer (SRE)C$180.2k – C$233.2k/year
ApplyView job
Kinaxis1 day ago

Cloud Engineer 2, Site Reliability Engineering

CA flagCanada OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers