
Senior Platform Engineer
Posted Aug 4

Posted Aug 4
This is a fully remote position, open to applicants in United States.
• Take ownership and enhance the platform foundation across cloud infrastructure, Kubernetes, infrastructure-as-code, secrets management, networking, access controls, CI/CD, observability, and production guardrails.
• Develop internal tools to support AI-enabled engineering workflows, including automation, repository and CI feedback loops, agent-ready development environments, and safeguards.
• Enhance logging, metrics, tracing, alerting, runbooks, and incident resolution processes.
• Strengthen production access by implementing least-privilege IAM, secure secret management, auditability, and controlled break-glass mechanisms.
• Set pragmatic platform standards that facilitate rapid team execution while minimizing infrastructure, reliability, and security liabilities.
• Over 5 years of experience in platform engineering, DevOps, SRE, infrastructure engineering, or backend-related cloud operations.
• Proven history of managing production systems where reliability, security, and developer efficiency are crucial.
• Practical experience with cloud infrastructure, Kubernetes, infrastructure-as-code, CI/CD, secrets management, access controls, and observability.
• Experience in developing internal developer tools, platform automation, or AI-assisted development workflows.
• Capability to design safe release processes with deployment gates, smoke tests, rollback paths, and defined ownership.
• Hands-on experience with supporting relational databases and managing production data changes.
• A security-focused approach to infrastructure, including least privilege, auditability, secret handling, and controlled access to production environments.
• Strong written communication skills for creating runbooks, deployment documentation, incident follow-ups, and engineering decision records.
• Familiarity with AWS or similar cloud platforms, managed Kubernetes, container registries, IAM, private networking, and secure cluster access.
• Experience with Terraform, OpenTofu, Terragrunt, or other infrastructure-as-code tools.
• Knowledge of GitHub Actions or comparable CI/CD systems, protected environments, federated identity, deployment gates, image pipelines, and smoke-test automation.
• Nice to have: familiarity with agent harness design, agent sandboxing, production model inference, GPU platforms, or self-hosted inference stacks.
• Base compensation ranging from $180K to $220K.
• Bonus offered.
Quantiphi
Agility Technologies Inc
American College of Education
First Due
Get handpicked remote jobs straight to your inbox weekly.