Remotery

Staff DevOps Engineer

Posted 8 hours ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Lead the technical strategy and architecture of Identity Digital’s cloud infrastructure.

• Design, develop, and manage a cloud-native platform for a multi-tenant managed service.

• Ensure production readiness for enterprise deployment, including SLOs, error budgets, runbooks, failover testing, and business continuity measures.

• Manage DNS production infrastructure, encompassing zones, DNSSEC, registrar/provider integrations, and failover mechanisms.

• Oversee signing and certificate infrastructure in collaboration with security engineering.

• Construct the delivery pipeline and CI enforcement boundary for an AI-assisted engineering team.

• Release signed versions with provenance and supply-chain controls.

• Integrate the platform with logging, metrics, and tracing capabilities.

• Automate infrastructure for audit-readiness, including logging, retention, access reviews, and evidence gathering.

• Track cloud expenditure and unit economics per tenant and workload.

• Evaluate designs, guide engineers, and assist in establishing hiring standards.

• Develop the on-call practice, act as incident commander, and facilitate blameless postmortems.

• Exemplify and advocate for Identity Digital’s core values.

• Perform additional duties as assigned.


⛳️ Requirements

• 10+ years of experience in DevOps, site reliability, or platform engineering, including principal or staff-level responsibilities.

• Bachelor’s degree in a relevant discipline or equivalent professional experience.

• Extensive experience with cloud-native architectures and services; AWS/GCP preferred.

• Strong expertise in Kubernetes, Terraform, and contemporary infrastructure-as-code methodologies.

• Practical experience with CI/CD tools such as GitHub Actions, GitLab CI, or ArgoCD.

• Proficient in scripting or programming languages such as Python, Go, or TypeScript/JavaScript.

• Experience in constructing observability stacks, including Prometheus, Grafana, ELK/EFK, or OpenTelemetry.

• Proficient knowledge in DNS operations, including zone management and DNSSEC.

• Comprehensive understanding of PKI, secrets management, and certificate workflows.

• Experience operating multi-tenant SaaS and distributed systems in production, including incident command roles.

• Proven history of enhancing teams’ operational standards through design reviews and mentorship.

• Willingness to travel as necessary.

• Capability to collaborate effectively across different time zones on a distributed team.

• Preferred: experience managing signing infrastructure or a certificate authority at scale.

• Preferred: expertise in supply-chain security involving SLSA, sigstore/cosign, reproducible builds, and provenance.

• Preferred: understanding of agentic AI systems and OWASP Agentic and NHI Top 10.

• Preferred: experience with infrastructure-side SOC 2 or equivalent audit programs.

• Preferred: experience transitioning a product from prototype to general availability.

• Preferred: contributions to open-source infrastructure or operations tooling.

• Must be able to pass a satisfactory background check.

• Work sponsorship may not be available now or in the future.

• Must be able to sit for extended periods and work at a computer.

• Must be able to lift up to 15 pounds at times.


🏝️ Benefits

• Generously subsidized medical, dental, and vision insurance.

• Company contributions to Health Savings Accounts.

• Company-paid life and disability insurance.

• Optional employee-paid supplemental life, accidental death and dismemberment, critical illness, and accident insurance.

• 401(k) plan with up to a 5% match.

• 15 days of paid vacation annually, increasing to 20 days after one year.

• 5 days of paid sick leave.

• 13 paid holidays.

• 20 weeks of paid parental leave for birthing parents.

• 12 weeks of paid parental leave for others.

• Tuition reimbursement for qualifying expenses.

• Discretionary and/or nondiscretionary bonuses.

• Long-term incentive plan.

• Remote work arrangement.

• Occasional travel for team gatherings and industry events.

People also viewed

Level Data3 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Level Data3 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data4 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data4 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Identity Digital Inc.8 hours ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$175k – $220k/year
ApplyView job
Abenis Consultores9 hours ago

DevOps Engineer – Azure, Inglés Avanzado

CL flagChile OnlyFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers