Remotery

SRE – CloudOps, Practice Architect II

atTEKsystemsRemoteUS flagIllinoisFull-timeDevOps & Site Reliability Engineer (SRE)SeniorLead$148.2k – $222.4k/year

Posted 5 hours ago

This is a fully remote position, open to applicants in Illinois.

📋 Description

• Take ownership of the complete technical architecture and deployment of cloud platforms on AWS and GCP.

• Design and implement production-grade reference architectures for multi-account AWS, multi-project GCP, Kubernetes, hybrid, and multi-cloud environments.

• Validate architectures via practical PoCs and intricate customer implementations.

• Spearhead platform design utilizing SRE principles, including SLIs, SLOs, error budgets, incident response, and post-incident reviews.

• Establish operational readiness standards, runbooks, and escalation models in collaboration with SRE and operations teams.

• Direct the integration of agentic systems with cloud platforms, Kubernetes, reliability engineering, operational controls, security, and governance.

• Provide mentorship to offshore teams developing agent-enabled platform capabilities.

• Construct infrastructure using Terraform and oversee Kubernetes provisioning, upgrades, and scaling.

• Guide CI/CD and GitOps-based platform delivery utilizing Helm, pipelines, and automation.

• Create reusable platform components, accelerators, and templates.

• Review and endorse designs related to networking, identity, security, ingress, and secrets.

• Lead and mentor distributed global teams and offshore members.

• Engage in code, design, and architecture reviews.

• Propel solutions from concept through reference architecture, implementation, and market-ready offering.

• Assist in pre-sales, solutioning, and technical validation.

• Facilitate quicker client onboarding through standardized patterns and pre-built assets.


⛳️ Requirements

• 10–12+ years of experience in cloud architecture, platform engineering, SRE, or distributed systems.

• Extensive hands-on knowledge in AWS, GCP, or hybrid production-scale environments.

• Demonstrated systems design and architecture expertise in enterprise settings.

• Profound understanding of agentic platform concepts, architectures, and their applicability in enterprise contexts.

• Advanced skills in performance tuning and troubleshooting across hybrid platforms.

• Strong foundation in networking, including VPCs, L3/L4, DNS, and load balancing.

• Solid knowledge of Kubernetes, encompassing EKS and GKE.

• Proficient in Terraform.

• Proficient in at least one programming or scripting language: Go, Python, or Bash.

• Experience in highly collaborative global delivery environments that include offshore teams.

• Experience managing and collaborating with large global teams.

• Ability to mentor, influence, and elevate technical standards without formal authority.

• Professional-level Cloud Solution Architect certification in AWS, Azure, or GCP is required.

• Excellent communication skills are essential.

• Client-facing experience is a must.

• Certifications in observability platforms are preferred.

• Terraform/IaC certifications are preferred.

• Hands-on experience in implementing agent-based or AI-enabled platforms is preferred.

• Experience in designing multi-tenant Kubernetes platforms is preferred.

• Strong observability experience with metrics, logs, tracing, Dynatrace, or Datadog is preferred.

• Experience in developing practice-level accelerators is preferred.

• Contributions to open-source or development of internal platform frameworks are preferred.

• Experience delivering solutions under tight deadlines with enterprise constraints is preferred.


🏝️ Benefits

• Medical, Dental, and Vision coverage.

• Critical Illness, Accident, and Hospital benefits.

• 401(k) Retirement Plan – Options for pre-tax and Roth post-tax contributions available.

• Life Insurance (Voluntary Life and AD&D for both employees and dependents).

• Short and Long-Term Disability coverage.

• Health Spending Account (HSA).

• Transportation Benefits.

• Employee Assistance Program.

• Time Off/Leave (PTO, Vacation, or Sick Leave).

• Additional earnings potential through incentive programs such as annual bonuses, profit sharing, etc.

People also viewed

TEKsystems5 hours ago

SRE CloudOps Practice Architect II

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data9 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Level Data9 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data9 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data9 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Identity Digital Inc.13 hours ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$175k – $220k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers