
Senior Platform Engineer
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in United States.
β’ Design and manage Gravie's shared engineering platform, which encompasses CI/CD pipelines, service templates, deployment tools, and shared infrastructure services.
β’ Take ownership of multi-account AWS infrastructure as code utilizing Terraform, CDK, or both.
β’ Collaborate with architecture to implement and maintain paved paths.
β’ Provide on-call support, incident response, postmortems, and systemic remediation for platform services.
β’ Establish default observability through metrics, logs, traces, SLOs, and actionable alerts.
β’ Manage the platform's security posture, addressing identity and permission boundaries, secrets, network and data isolation, as well as remediation in HIPAA and SOC 2 environments.
β’ Enhance cloud cost attribution and visibility via tagging and waste reduction initiatives.
β’ Employ and integrate AI agents for code reviews, alert triage, migrations, and repetitive remediation tasks.
β’ Automate platform operations and evaluate the outcomes.
β’ Maintain documentation and facilitate developer onboarding.
β’ Promote adoption, establish standards, and oversee the migration of existing services.
β’ Conduct code reviews, mentor engineers, and assist teams in effectively utilizing AI tools.
β’ A minimum of five years of experience in software or infrastructure engineering, including ownership of production systems that you have also operated.
β’ Profound AWS expertise, covering IAM policy evaluation, VPC routing and DNS, and container or serverless runtime behavior.
β’ Hands-on management of a multi-account AWS environment, including identity boundaries, networking, managed data services, and governance of costs and tagging.
β’ Experience with infrastructure as code using Terraform, CDK, or both, with skills in review, testing, and blast-radius analysis.
β’ Strong background in CI/CD engineering, including shared or templated pipelines that are utilized by multiple teams.
β’ Experience working within or closely with a platform, infrastructure, or developer experience team.
β’ Familiarity with AI tooling such as Claude Code for complex multi-step engineering tasks.
β’ Background in production operations: on-call duties, incident response, postmortems, and SLO considerations.
β’ Proficient programming skills in Python, Go, TypeScript, or a JVM language.
β’ Sound judgment regarding security, secrets, access control, and data management.
β’ Excellent written communication skills.
β’ Experience in healthcare or other regulated domains, with knowledge of HIPAA and SOC 2 obligations preferred.
β’ Preferred experience in a venture-backed high-growth company.
β’ Experience with GitLab CI at scale preferred.
β’ Familiarity with Datadog, ECS or Kubernetes, AI tooling development, Renovate, DORA, Team Topologies, and platform decommissioning is preferred.
β’ Authorization to work in the United States is required; visa sponsorship inquiries are included in the application.
β’ Equity
β’ Standard health and wellness benefits
β’ Coverage for alternative medicine
β’ Flexible PTO
β’ Up to 16 weeks of paid parental leave
β’ Paid holidays
β’ 401k program
β’ Transportation perks
β’ Education reimbursement
β’ 2 days of paid paw-ternity leave
Planet Technologies
Northrop Grumman
Credit Acceptance
Planet Technologies
Get handpicked remote jobs straight to your inbox weekly.