Manager, DevOps

Posted 3 days ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Take ownership of the DevOps roadmap and delivery commitments for all Supplier.io environments.

• Oversee, mentor, develop, recruit, and onboard the DevOps team.

• Implement and uphold processes for change management, release management, environment promotion, incident response, on-call duties, documentation, and runbooks.

• Transition legacy infrastructure to a version-controlled and documented state.

• Manage disaster recovery and business continuity strategies, including recovery time and recovery point objectives.

• Establish sprint planning, estimation, and delivery commitments.

• Define and promote observability standards for monitoring, logging, alerting, and Application Performance Management (APM).

• Oversee cloud cost optimization efforts, including committed-use discounts, right-sizing, and minimizing logging and storage expenses.

• Enhance security and compliance practices, focusing on least-privilege access, access reviews, key rotation, certificate automation, and vulnerability remediation.

• Develop and evaluate CI/CD pipelines, Infrastructure as Code, and Kubernetes workloads.

• Collaborate with Data Engineering on Snowflake administration, Airflow infrastructure, and secrets and connection management.

• Align infrastructure decisions with the STAC and ensure compliance with company-wide architectural standards.

• Participate in the on-call rotation.

• Foster a culture of automation, accountability, and continuous improvement.

• Collaborate with software development, data engineering, AI, and platform teams.

• Assess emerging tools and practices to determine whether to build, buy, or adopt.


⛳️ Requirements

• Over 10 years of experience in DevOps, cloud engineering, or infrastructure roles with increasing responsibilities.

• At least 3 years of experience leading or managing engineers, including hiring, coaching, and performance evaluation.

• Proven success in establishing change and release management, incident response, planning and estimation, and documentation processes.

• Experience in managing legacy or undocumented infrastructure.

• Background in implementing Infrastructure as Code and disaster recovery planning.

• Strong practical knowledge of CI/CD platforms such as Azure DevOps, GitHub, Jenkins, or similar tools.

• Proficient in Infrastructure as Code tools, including Terraform, Ansible, and Helm.

• Experience with Docker and Kubernetes technologies.

• Extensive experience with Google Cloud Platform.

• Skilled in defining and executing observability strategies across various environments.

• Experience managing cloud expenditures, including committed-use or reservation strategies, right-sizing, and cost accountability.

• Exceptional communication abilities.

• Judicious in balancing operations, project delivery, and investments in automation.

• Experience in building or managing distributed teams, including offshore execution teams (preferred).

• Background in consolidating infrastructure and operations across acquired platforms (preferred).

• Familiarity with security automation, compliance frameworks, and policy-as-code (preferred).

• Understanding of CI/CD for data pipelines and AI/MLOps practices (preferred).

• Candidates must meet in-country work authorization requirements.

• Supplier.io cannot sponsor work visas for positions in the US.


🏝️ Benefits

• Equal Employment Opportunity employer.

• Reasonable accommodations available for the application or interview process.

People also viewed

Horizon3.ai1 day ago

Staff Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$199.8k – $270k/year
ApplyView job
CLOUD MANTA GmbH1 day ago

Senior DevOps Engineer, Containers & Private Cloud

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€70k – €80k/year
ApplyView job
Stefanini LATAM1 day ago

Senior DevOps

AR flagArgentina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Akamai Technologies1 day ago

Principal Site Reliability Engineer – Lead

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
PingWind Inc. (SDVOSB)1 day ago

DevSecOps Engineer

US flagAlabama, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ad Hoc LLC1 day ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers