
Senior Site Reliability Engineer
Posted 12 hours ago

Posted 12 hours ago
This is a fully remote position, open to applicants in United States.
• Design and enhance both new and existing systems to boost performance, reliability, and scalability.
• Develop, implement, and refine CI/CD pipelines.
• Oversee, create, design, and deploy microservice and containerized applications.
• Establish security measures in distributed systems and agents.
• Automate deployment processes and configurations across various platforms.
• Create scalable automation for implementing observability.
• Recognize opportunities for observability and process enhancements.
• Standardize alerts and notifications while responding to monitoring tools.
• Integrate observability into daily application operations.
• Participate in post-mortems, conduct root cause analysis, and follow up on corrective action items.
• Advocate for DevOps best practices and Agile/Scrum methodologies.
• Contribute to hybrid-cloud production containerization services.
• Design and enforce standards, policies, and procedures for automation and integrations.
• Acquire knowledge of toolsets and implement features to optimize operations.
• Ensure that production security systems meet uptime requirements and stay up to date.
• Bachelor’s Degree with 7 years of experience; Master’s Degree with 6 years of experience; or PhD with 2 years of experience.
• Consider security best practices as an essential requirement.
• Proficient in AWS, GCP, or Azure cloud platform administration.
• Familiar with the pillars of observability.
• Experienced in high-scale environments and distributed architectures.
• Knowledgeable in Agile and DevOps methodologies.
• Familiar with CI/CD tools like GitHub Actions, Bamboo, Jenkins, or Azure DevOps.
• Comfortable working with Docker workloads and Kubernetes / Amazon ECS.
• Capable of working independently as well as collaboratively.
• Proficient in managing both Linux and Windows environments.
• Preferred: Experience with SPIRE/SPIFFE, Terraform/Crossplane, development tools and scripting languages, MCP Servers, database management systems, AWS Cloud Practitioner / Azure AZ-900, serverless architecture, containerized applications, data management and pipeline technologies, Agile teams, OpenTelemetry, Prometheus/Grafana, Kubernetes distributed platforms, GitOps, and Infrastructure as Code.
• Open to remote work from anywhere in the U.S.
• Willing to travel up to 10% of the time.
• Paid time off including vacation, holidays, and sick leave.
• Medical, dental, and vision insurance coverage.
• 401(k) retirement plan.
• Short-term incentive programs.
• Eligibility for remote work.
• Opportunities for travel up to 10%.
Akamai Technologies
BeyondTrust
Cencora
PrizePicks
Get handpicked remote jobs straight to your inbox weekly.