Staff Engineer – Site Reliability Engineering

Posted Sep 18

This is a fully remote position, open to applicants in Pennsylvania.

📋 Description

• Create and implement a scalable Site Reliability Engineering (SRE) ecosystem adhering to SRE and DevSecOps best practices.

• Develop reusable TypeScript scaffolding libraries tailored for cloud-native components.

• Build and improve solutions utilizing AWS, EKS, Kubernetes, and Infrastructure as Code methodologies.

• Promote automation, reliability, scalability, and operational excellence throughout microservices.

• Define and execute SRE and DevOps best practices across various applications.

• Collaborate with technology, product, operations, and functional teams on SRE initiatives.

• Analyze business requirements and assess their impact on applications and cloud systems.

• Establish standardized and automated onboarding pathways for applications onto the SRE platform.

• Assess emerging technologies and formulate strategies for cloud and SRE adoption.

• Implement CI/CD, observability, monitoring, and service mesh capabilities.

• Identify and resolve reliability, scalability, and operational challenges.

• Offer technical guidance and support to distributed engineering teams.


⛳️ Requirements

• Minimum of 5.5 years of total experience.

• Strong proficiency in TypeScript development, including coding and design patterns.

• Practical experience with AWS cloud services and cloud-native technologies.

• Familiarity with AWS CDK, Terraform, and Infrastructure as Code (IaC).

• Knowledge of Kubernetes, Docker, and Amazon EKS.

• Experience in SRE, DevOps, scalability, reliability, and cloud automation.

• Proficiency with CI/CD tools like Jenkins and Git.

• Understanding of observability and monitoring tools such as CloudWatch, Splunk, and Dynatrace.

• Knowledge of service mesh technologies, including Istio.

• Experience in developing reusable scaffolding libraries and cloud-native components.

• Capability to analyze application and infrastructure dependencies across microservices environments.

• Experience working with distributed teams across various time zones.

• Excellent communication, presentation, and stakeholder collaboration skills.

• Bachelor’s or master’s degree in computer science, Information Technology, or a related field.


🏝️ Benefits

• Employees have the flexibility to work remotely.

People also viewed

Horizon3.ai1 day ago

Staff Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$199.8k – $270k/year
ApplyView job
CLOUD MANTA GmbH1 day ago

Senior DevOps Engineer, Containers & Private Cloud

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€70k – €80k/year
ApplyView job
Stefanini LATAM1 day ago

Senior DevOps

AR flagArgentina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Akamai Technologies1 day ago

Principal Site Reliability Engineer – Lead

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
PingWind Inc. (SDVOSB)1 day ago

DevSecOps Engineer

US flagAlabama, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ad Hoc LLC1 day ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers