Site Reliability Engineering Lead

Posted 8 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Lead a small, senior Site Reliability Engineering (SRE) team, providing strategic planning, clarity, and opportunities for career advancement.

• Direct the team's efforts towards impactful reliability initiatives rather than manual, ticket-based responses.

• Develop production-level code and create automation, tooling, and best practices.

• Contribute to the team's codebase and engage in design and code review processes.

• Establish service ownership, implement alerting systems, create dashboards, and ensure safe deployment and rollback procedures.

• Conduct regular reviews of service health for the systems under support.

• Take charge of incident command when necessary and drive confirmed remediation efforts.

• Create self-service incident management tools and automate runbooks.

• Design a sustainable on-call coverage strategy through effective staffing, handoffs, and automation.

• Collaborate with service teams and engineering leaders to enhance reliability benchmarks.

• Report directly to the Head of Platform and Data Engineering.


⛳️ Requirements

• A solid foundation in software engineering.

• Ability to write production-quality code in contemporary programming languages such as Python or Go.

• Comfortable being evaluated on coding skills.

• In-depth knowledge of reliability for production systems and services at scale.

• Ideally, experience in a consumer, fintech, or other high-availability sectors.

• A background established in a high-caliber, high-scale engineering organization with a robust engineering culture.

• Recent experience in people leadership; managed a small engineering team within the last few years.

• Practical experience with observability, alerting, safe deployment practices, automation, and incident management.

• Effective and composed communication skills with team members, partner teams, and during live incidents.

• Authorized to work legally in the United States.

• Engineering professional experience is required on the application form, with options including 0–3, 3–7, 7–10, or 10+ years.


🏝️ Benefits

• Stock options.

• Health benefits starting from Day 1.

• 401(k) plan with company match.

• Fully remote work opportunity (within the US).

• Flexible time off (FTO).

• Opportunities for professional growth.

• A high-growth, mission-driven, and inclusive culture.

People also viewed

Arista Networks9 hours ago

FedRAMP Site Reliability Engineer – CloudVision

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$101k – $161k/year
ApplyView job
Octus9 hours ago

Lead DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$175k – $225k/year
ApplyView job
Tandem Diabetes Care9 hours ago

Principal Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$165k – $185k/year
ApplyView job
TechnologyAdvice12 hours ago

Senior DevOps Engineer – Contract

IN flagIndia OnlyFreelanceDevOps & Site Reliability Engineer (SRE)₹1,500 – ₹2,000/hour
ApplyView job
Bixal12 hours ago

Director of DevSecOps

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$165k – $195k/year
ApplyView job
Sphera12 hours ago

Cybersecurity Engineer – DevOps

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$112k – $178k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers