
Senior Platform Engineer
Posted 19 hours ago

Posted 19 hours ago
This is a fully remote position, open to applicants in Texas.
• Take ownership and enhance AWS infrastructure encompassing compute, networking, storage, and managed services.
• Create and uphold infrastructure that supports high availability, reliable performance, and financial accuracy.
• Direct platform architectural decisions, including service migrations and runtime modifications such as transitioning from Redis to Valkey and from EKS to ECS/Fargate.
• Ensure that infrastructure selections align with reliability, cost efficiency, and operational simplicity.
• Design and sustain secure, repeatable, and observable deployment pipelines.
• Manage reliability through capacity planning, failure modeling, and controlled change management.
• Spearhead incident response efforts and conduct root-cause analysis for infrastructure failures.
• Engage in on-call rotations and enhance operational ergonomics.
• Develop and uphold observability through metrics, logs, tracing, and alerting mechanisms.
• Safeguard AWS resources, IAM policies, secrets management, and network boundaries.
• Recognize and mitigate infrastructure risks associated with scale, cost, and security.
• Collaborate with application engineers to address platform constraints and capabilities.
• Propel infrastructure modifications through direct implementation.
• Establish standards for infrastructure, deployment, and operations as the team expands.
• Mentor platform engineers to enhance the organization's operational maturity.
• Report directly to the Director of Engineering.
• 8+ years of experience in building and operating production infrastructure within cloud environments.
• Extensive experience with AWS core services, including EC2, ECS/EKS, VPC, IAM, RDS, ElastiCache, ALB/NLB, and CloudWatch.
• Strong comprehension of containerized workloads and the trade-offs involved in orchestration.
• Proven track record in designing systems that ensure high availability, fault tolerance, and managed failure.
• Practical experience with infrastructure as code tools, such as Terraform or CloudFormation.
• Demonstrated capability to plan and execute infrastructure migrations safely.
• Experience troubleshooting production incidents related to networking, scaling, or service degradation.
• Preferred: Experience managing infrastructure for financial systems or other high-reliability domains.
• Preferred: Direct experience in scaling infrastructure during periods of rapid growth.
• Preferred: Strong views on observability and operational hygiene shaped by past failures.
• Preferred: Experience in simplifying, decommissioning, or rearchitecting overly complex infrastructure.
• Preferred: Comfort in collaborating closely with a Rails-based application stack while remaining tool-agnostic.
• Generous PTO and company holiday policy.
• Company-paid Short Term Disability benefits.
• 100% employer-covered health and dental insurance for directly employed individuals on the specified plan.
• Higher-tier healthcare coverage available at the employee’s additional expense.
• Dependent healthcare coverage available at the employee’s cost.
• Vision plan available at the employee’s additional cost.
• Child Care Benefits.
• Generous parental leave policy.
The PSSG
Coinbase
Prescryptive Health, Inc.
Get handpicked remote jobs straight to your inbox weekly.