DevOps / Site Reliability Engineer

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Set up, manage, and enhance Bitbucket pipelines for application deployment to both staging and production environments.

• Collaborate with the Cloud Security Engineer to boost the speed, reliability, and security of the CI pipeline.

• Support developers and QA teams during deployment processes.

• Utilize Docker and AWS ECR for building containers and managing deployment workflows.

• Analyze and resolve system issues flagged by Sentry, New Relic, and CloudWatch.

• Oversee application performance, identify bottlenecks, and suggest solutions.

• Address production and staging issues, including database latency, unresponsive resources, or job failures.

• Provide maintenance and support for non-production environments utilized by developers and QA.

• Enhance and sustain AWS infrastructure along with Terraform resources.

• Execute updates and upgrades to AWS services to guarantee reliability and scalability.

• Collaborate with engineers to design systems that are scalable, observable, and resilient.

• Work alongside the cloud security engineer to ensure secure configurations in CI/CD, AWS, and containerized workloads.

• Contribute to workflow enhancements, automation, and monitoring strategies.

• Utilize AI to automate monitoring and diagnostic processes.


⛳️ Requirements

• A minimum of 3 years of experience in DevOps, SRE, or similar engineering roles.

• Extensive experience in configuring CI/CD pipelines (Bitbucket Pipelines, GitHub Actions, or equivalent).

• Proficient in configuring, debugging, and deploying PHP applications.

• Practical experience with Docker and AWS ECR for container builds and deployments.

• Strong familiarity with AWS services (EC2, RDS, ECS, Lambda, etc.) and Terraform for infrastructure as code.

• Knowledge of monitoring and observability tools such as New Relic, Sentry, CloudWatch, or similar.

• Solid troubleshooting capabilities for diagnosing performance issues in databases, applications, and distributed systems.

• Experience with contemporary software development workflows (agile teams, code reviews, branching strategies).

• Robust scripting and automation skills (Bash, Python, or similar).

• Exceptional communication skills and a collaborative approach.

• Interest in using AI agents to automate monitoring and diagnosis workflows.


🏝️ Benefits

• Competitive salary and performance-based bonuses.

• Comprehensive health, dental, and vision insurance.

• Opportunities for professional development and continuous learning.

• Flexible work hours and remote work options.

• A collaborative and dynamic work environment.

People also viewed

Horizon3.ai1 day ago

Staff Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$199.8k – $270k/year
ApplyView job
CLOUD MANTA GmbH1 day ago

Senior DevOps Engineer, Containers & Private Cloud

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€70k – €80k/year
ApplyView job
Stefanini LATAM1 day ago

Senior DevOps

AR flagArgentina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Akamai Technologies1 day ago

Principal Site Reliability Engineer – Lead

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
PingWind Inc. (SDVOSB)1 day ago

DevSecOps Engineer

US flagAlabama, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ad Hoc LLC1 day ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers