Senior Site Reliability Engineer

Posted 12 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Design and enhance both new and existing systems to boost performance, reliability, and scalability.

• Develop, implement, and refine CI/CD pipelines.

• Oversee, create, design, and deploy microservice and containerized applications.

• Establish security measures in distributed systems and agents.

• Automate deployment processes and configurations across various platforms.

• Create scalable automation for implementing observability.

• Recognize opportunities for observability and process enhancements.

• Standardize alerts and notifications while responding to monitoring tools.

• Integrate observability into daily application operations.

• Participate in post-mortems, conduct root cause analysis, and follow up on corrective action items.

• Advocate for DevOps best practices and Agile/Scrum methodologies.

• Contribute to hybrid-cloud production containerization services.

• Design and enforce standards, policies, and procedures for automation and integrations.

• Acquire knowledge of toolsets and implement features to optimize operations.

• Ensure that production security systems meet uptime requirements and stay up to date.


⛳️ Requirements

• Bachelor’s Degree with 7 years of experience; Master’s Degree with 6 years of experience; or PhD with 2 years of experience.

• Consider security best practices as an essential requirement.

• Proficient in AWS, GCP, or Azure cloud platform administration.

• Familiar with the pillars of observability.

• Experienced in high-scale environments and distributed architectures.

• Knowledgeable in Agile and DevOps methodologies.

• Familiar with CI/CD tools like GitHub Actions, Bamboo, Jenkins, or Azure DevOps.

• Comfortable working with Docker workloads and Kubernetes / Amazon ECS.

• Capable of working independently as well as collaboratively.

• Proficient in managing both Linux and Windows environments.

• Preferred: Experience with SPIRE/SPIFFE, Terraform/Crossplane, development tools and scripting languages, MCP Servers, database management systems, AWS Cloud Practitioner / Azure AZ-900, serverless architecture, containerized applications, data management and pipeline technologies, Agile teams, OpenTelemetry, Prometheus/Grafana, Kubernetes distributed platforms, GitOps, and Infrastructure as Code.

• Open to remote work from anywhere in the U.S.

• Willing to travel up to 10% of the time.


🏝️ Benefits

• Paid time off including vacation, holidays, and sick leave.

• Medical, dental, and vision insurance coverage.

• 401(k) retirement plan.

• Short-term incentive programs.

• Eligibility for remote work.

• Opportunities for travel up to 10%.

People also viewed

Akamai Technologies6 hours ago

Senior Site Reliability Engineer

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
BeyondTrust6 hours ago

Senior Site Reliability Engineer

US flagUnited States, +1 more countryFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Cencora6 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
PrizePicks9 hours ago

Manager, Database Reliability Engineering

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$200k – $220k/year
ApplyView job
onXmaps, Inc.13 hours ago

Site Reliability Engineer III

US flagColorado, +6 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $153k/year
ApplyView job
hims & hers14 hours ago

Senior Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$140k – $165k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers