Remotery

SRE CloudOps Practice Architect II

atTEKsystemsRemoteUS flagTexasFull-timeDevOps & Site Reliability Engineer (SRE)SeniorLead$148.2k – $222.4k/year

Posted 5 hours ago

This is a fully remote position, open to applicants in Texas.

📋 Description

• Take ownership of the complete technical architecture and execution of cloud platforms on AWS and GCP.

• Create and implement production-ready reference architectures for multi-account AWS, multi-project GCP, Kubernetes, hybrid, and multi-cloud settings.

• Assess architectures, develop proofs of concept, and assist in complex customer implementations with a hands-on approach.

• Direct platform design following SRE principles, which include SLIs, SLOs, error budgets, incident response, and post-incident evaluations.

• Establish operational readiness standards, runbooks, and escalation models in collaboration with SRE and operations teams.

• Facilitate the integration of agentic systems with cloud platforms, Kubernetes, reliability engineering, operational controls, security, and governance.

• Provide mentorship to offshore teams working on agent-enabled platform capabilities.

• Develop Terraform infrastructure, oversee Kubernetes lifecycles, and deliver platforms via CI/CD and GitOps.

• Advise on reusable platform components, accelerators, and templates.

• Review and authorize shared platform service designs encompassing networking, identity, security, ingress, and secrets.

• Lead and mentor global and offshore teams.

• Engage in code, design, and architecture evaluations.

• Connect architecture, delivery, and operations to create executable client solutions.

• Propel solution incubation from concept to reference architecture, implementation, and market-ready offerings.

• Facilitate client onboarding with standardized patterns and pre-built assets.

• Assist in pre-sales, solution development, and technical validation while maintaining a hands-on approach.


⛳️ Requirements

• 10–12+ years of experience in cloud architecture, platform engineering, SRE, or distributed systems.

• Extensive hands-on knowledge of AWS, GCP, or hybrid production-scale environments.

• Proven experience in systems design and architecture in real-world enterprise scenarios.

• Strong grasp of agentic platform concepts, architectures, and their applicability in enterprises.

• Advanced skills in performance tuning and troubleshooting across hybrid platforms.

• Solid networking fundamentals, including VPCs, L3/L4, DNS, and load balancing.

• Comprehensive knowledge of Kubernetes, including EKS and GKE.

• Proficient in Terraform and at least one programming language such as Go, Python, or Bash.

• Experience working in highly collaborative global delivery settings with offshore teams.

• Proven ability to manage and collaborate with a large global team.

• Capacity to mentor, influence, and elevate technical standards without formal authority.

• Professional-level Cloud Solution Architect certification from AWS, Azure, or GCP is required.

• Strong communication skills are essential as this is a client-facing role.

• Certifications in observability platforms such as Dynatrace or Datadog are preferred.

• Terraform/IaC certifications are preferred.

• Hands-on experience with agent-based or AI-enabled platforms is preferred.

• Experience in designing multi-tenant Kubernetes platforms is preferred.

• Strong observability experience with metrics, logs, and tracing at scale is preferred.

• Experience in building practice-level accelerators is preferred.

• Contributions to open source or development of internal platform frameworks are preferred.

• Ability to deliver solutions under tight timelines within enterprise constraints is preferred.


🏝️ Benefits

• Medical, Dental, and Vision coverage.

• Critical Illness, Accident, and Hospital benefits.

• 401(k) Retirement Plan – Options for pre-tax and Roth post-tax contributions available.

• Life Insurance (Voluntary Life and AD&D for employee and dependents).

• Short and Long-Term Disability coverage.

• Health Spending Account (HSA).

• Transportation Benefits.

• Employee Assistance Program.

• Time Off/Leave (PTO, Vacation, or Sick Leave).

• Additional earnings may be available through annual bonuses, profit sharing, and other incentive programs.

People also viewed

TEKsystems5 hours ago

SRE – CloudOps, Practice Architect II

US flagIllinois OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data9 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Level Data9 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data9 hours ago

DevOps Engineer II

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$95k – $110k/year
ApplyView job
Level Data9 hours ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job
Identity Digital Inc.13 hours ago

Staff DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$175k – $220k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers