
Senior Platform Engineer
Posted Jul 15

Posted Jul 15
This is a fully remote position, open to applicants in United States.
• As a Senior Platform Engineer, you will be a key member of a pioneering team with a clear, three-part mission: to completely rearchitect and codify our cloud infrastructure on AWS and Azure, establish an exceptional SRE and observability practice, and create an Internal Developer Platform featuring golden paths that facilitate swift, secure, and self-service software shipping.
• Cloud Modernization — Rearchitect & Codify: Transition all existing AWS and Azure infrastructures to OpenTofu/Terraform and Ansible; set module standards, manage remote state, and implement GitOps-based plan/apply pipelines — ensuring no unmanaged resources.
• Implement policy-as-code (OPA/Conftest, AWS SCPs, Azure Policy) to enforce security, tagging, and compliance guardrails at the platform level — integrating governance rather than adding it later.
• Develop and maintain reusable Terraform modules for compute (EKS, AKS, EC2), networking, storage, databases, and identity, serving as shared building blocks for all engineering teams.
• Design and implement a comprehensive observability stack: metrics (Prometheus/Datadog), logs (Loki/OpenSearch), traces (Tempo/Datadog APM), and dashboards (Grafana) — instrumented end-to-end using OpenTelemetry.
• Define SLIs and SLOs for all platform shared services and critical applications; create error budget dashboards and burn-rate alerts — focusing on symptoms rather than raw metrics.
• Deploy and manage a developer portal (Backstage, GitHub, or equivalent) as the single access point: service catalog, scaffolding templates, runbooks, API documentation, and on-call ownership all consolidated in one location.
• At least 5 years of experience in platform, infrastructure, or DevOps engineering with direct production ownership on AWS and/or Azure.
• Extensive proficiency in OpenTofu/Terraform: module authoring, state management, workspace strategy, remote backends, and CI/CD integration; experience with Terramate is a plus.
• Strong operational knowledge of Kubernetes: managing EKS and/or AKS clusters, Helm, admission controllers, RBAC, network policies, and autoscaling.
• Hands-on experience with observability tools such as Prometheus, Grafana, Loki, Tempo, Datadog, or OpenTelemetry — including SLI/SLO definition and alert engineering.
• Experience with CI/CD platforms: authoring GitHub Actions pipelines, designing reusable workflows, and managing container build/scan pipelines.
• Familiarity with GitOps: using ArgoCD or Flux for Kubernetes continuous delivery; knowledge of progressive delivery patterns (canary, blue-green) is a strong advantage.
• Experience with IDP: working with Backstage or equivalent developer portals, GitHub, designing scaffolding templates, service catalogs, or self-service provisioning tools.
• A security-first approach: expertise in policy-as-code, IaC scanning, secrets management, container hardening, and shift-left security practices.
• Excellent communication and documentation skills; capable of presenting architectural decisions to engineering colleagues and leadership.
• Comprehensive benefits package including health, dental, and vision insurance.
• 401(k) plan with company matching.
• Generous paid time off to promote your well-being.
• Flexible work environment options, whether remote, hybrid, or in-office.
Arctiq
Cisco
Prove
Hello Heart
Get handpicked remote jobs straight to your inbox weekly.