DevOps Team Lead

Posted Aug 25

This is a fully remote position, open to applicants in Wisconsin.

📋 Description

• Take ownership of the delivery, reliability, and daily operations of the AWS cloud platform.

• Design, construct, evaluate, and manage Infrastructure-as-Code across a multi-account AWS environment.

• Lead and enhance infrastructure CI/CD pipelines with automated validation, security, compliance measures, and keyless cloud authentication.

• Build and sustain AWS networking, containers, databases, storage, messaging, systems management, backup, monitoring, and cost controls.

• Establish engineering standards for IaC, IAM, tagging, state management, versioning, security, documentation, and operational readiness.

• Convert product, engineering, security, and data requirements into actionable milestones, estimates, dependencies, and acceptance criteria.

• Manage and optimize the DevOps backlog while facilitating Scrum ceremonies.

• Enhance reliability through observability, SLOs, alerting, runbooks, disaster recovery, backup and restore, and reduction of operational toil.

• Participate in a rotating night and weekend on-call schedule and provide technical leadership during incidents.

• Collaborate with security on identity management, least privilege policies, network segmentation, encryption, secrets management, audit logging, and compliance.

• Lead efforts in cloud cost visibility and optimization, including tagging, right-sizing, commitment planning, anomaly response, and multi-region cost analysis.

• Advance the platform towards globally distributed workloads, focusing on regional rollout, data residency, replication, latency-aware routing, and cross-region failover.

• Utilize AI tools in engineering workflows and support infrastructure for emerging AI capabilities.

• Set the technical direction with architecture and mentor engineers through reviews, pairing, design discussions, documentation, and shared ownership.


⛳️ Requirements

• Extensive experience in software engineering, infrastructure, or DevOps, with a proven record of operating at a senior or technical lead level in complex cloud environments.

• Expert-level, hands-on experience with Terraform or equivalent Infrastructure as Code.

• Experience with IaC at scale, encompassing reusable module design and versioning, state architecture, safe refactoring, upgrades, drift detection and remediation, and transitioning legacy infrastructure under IaC management.

• Profound hands-on AWS experience across compute, storage, networking, IAM, and security, preferably within a multi-account AWS Organization.

• Demonstrated ownership of infrastructure CI/CD pipelines, including automated policy, validation, and security controls.

• Experience in designing and operating containerized or distributed production workloads with suitable networking, security, and observability.

• Proficiency in scripting and a history of automating operational tasks.

• Proven ownership of production operations, including on-call duties, incident response, runbook development, and independently troubleshooting and resolving platform issues.

• Experience in leading Agile/Scrum practices, including backlog ownership, refinement, estimation, planning, and stakeholder negotiations.

• Ability to translate ambiguous needs into actionable technical plans that consider requirements, dependencies, risks, milestones, and estimates.

• Strong grasp of cloud security, compliance, and risk management, including least privilege, zero trust, encryption, and exception management.

• Excellent written communication, documentation, and stakeholder engagement abilities.

• Self-directed and collaborative approach focused on system improvements.

• Willingness and capability to participate in a rotating night and weekend on-call schedule and manage production incidents independently.

• Preferred: Experience with Microsoft Azure.

• Preferred: Experience with globally distributed or multi-region applications.

• Preferred: Knowledge of monorepo IaC, policy-as-code, IaC security tooling, or automated code-quality platforms.

• Preferred: Experience in FinOps.

• Preferred: Familiarity with SRE practices.

• Preferred: Background in analytics, data engineering, database, or integration workloads.

• Preferred: AWS certification.

• Preferred: Experience with hybrid on-premises connectivity or improvements in the software development lifecycle.

• Preferred: Experience in AI/ML infrastructure.


🏝️ Benefits

• Competitive salary and performance-based bonuses.

• Comprehensive health, dental, and vision insurance plans.

• Generous paid time off and flexible work schedules.

• Opportunities for professional development and continuing education.

• Supportive and inclusive work environment.

People also viewed

Cisco13 hours ago

Site Reliability Engineer – FEDRAMP

US flagAlabama, +25 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$128.6k – $184.9k/year
ApplyView job
Stefanini Brasil13 hours ago

Integration Architect – DevSecOps, OpenShift, Kubernetes

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Study Now15 hours ago

DevOps Engineer

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)£1,170 – £1,950/month
ApplyView job
Spring Financial15 hours ago

DevOps Engineer II – Contract

MX flagMexico OnlyFreelanceDevOps & Site Reliability Engineer (SRE)$612k – $857k/year
ApplyView job
Endeavor15 hours ago

DevOps Lead

US flagConnecticut, +3 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$112.5k – $150k/year
ApplyView job
MOXFIVE16 hours ago

Senior DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$110k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers