DevOps/Site Reliability Engineer

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Set up, manage, and optimize Bitbucket pipelines for both staging and production environments.

• Enhance CI pipeline efficiency, dependability, and security alongside the Cloud Security Engineer.

• Support developers and QA teams during deployment processes.

• Utilize Docker and AWS ECR for container creation and deployment workflows.

• Investigate system issues reported by Sentry, New Relic, and CloudWatch.

• Monitor application performance, detect bottlenecks, and suggest improvements.

• Address production and staging challenges, such as database latency, unresponsive resources, and job failures.

• Maintain and provide support for non-production environments utilized by developers and QA teams.

• Oversee and enhance AWS infrastructure and Terraform resources.

• Update and upgrade AWS services to maintain reliability and scalability.

• Collaborate with engineers to design systems that are scalable, observable, and resilient.

• Ensure secure configurations within CI/CD, AWS, and containerized environments.

• Contribute to enhancements in workflows, automation, and monitoring strategies.

• Utilize AI technologies to automate monitoring and diagnostic processes.


⛳️ Requirements

• A minimum of 3 years of experience in DevOps, SRE, or related engineering positions.

• Extensive experience in configuring CI/CD pipelines, including Bitbucket Pipelines, GitHub Actions, or equivalent.

• Proficiency in configuring, debugging, and deploying PHP applications.

• Practical experience with Docker and AWS ECR for container builds and deployments.

• Extensive knowledge of AWS services such as EC2, RDS, ECS, and Lambda.

• Strong expertise in Terraform for infrastructure as code.

• Familiarity with monitoring and observability tools such as New Relic, Sentry, and CloudWatch.

• Excellent troubleshooting abilities for performance issues in databases, applications, and distributed systems.

• Experience with contemporary software development workflows, including agile methodologies, code reviews, and branching strategies.

• Strong scripting and automation capabilities using Bash, Python, or similar languages.

• Exceptional communication skills and a team-oriented mindset.

• Interest in utilizing AI agents to automate monitoring and diagnostic workflows.


🏝️ Benefits

• Flexible remote work options.

• Opportunity to engage with groundbreaking tools in mobile and web applications, computer vision, and LLMs.

• Collaboration with distributed teams throughout the United States.

People also viewed

Horizon3.ai1 day ago

Staff Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$199.8k – $270k/year
ApplyView job
CLOUD MANTA GmbH1 day ago

Senior DevOps Engineer, Containers & Private Cloud

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€70k – €80k/year
ApplyView job
Stefanini LATAM1 day ago

Senior DevOps

AR flagArgentina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Akamai Technologies1 day ago

Principal Site Reliability Engineer – Lead

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
PingWind Inc. (SDVOSB)1 day ago

DevSecOps Engineer

US flagAlabama, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ad Hoc LLC1 day ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers