
Platform Engineer
Posted 18 hours ago

Posted 18 hours ago
This is a fully remote position, open to applicants in United States.
• Develop, manage, and enhance AWS infrastructure encompassing compute, networking, storage, and managed services.
• Design and uphold infrastructure that ensures high availability, reliable performance, and financial accuracy.
• Participate in platform architectural decisions, including service migrations and runtime adjustments such as transitioning from Redis to Valkey and from EKS to ECS/Fargate.
• Create and sustain secure, repeatable, and observable deployment pipelines.
• Engage in capacity planning, failure modeling, controlled change management, incident response, and root-cause analysis.
• Be part of on-call rotations and enhance operational ergonomics.
• Develop and maintain observability across infrastructure and services, including metrics, logs, tracing, and alerting.
• Assist in securing AWS resources, IAM policies, secrets management, and network boundaries.
• Recognize and mitigate infrastructure risks related to scaling, costs, and security.
• Collaborate with application engineers on platform limitations and capabilities.
• Implement infrastructure modifications both independently and collaboratively.
• Contribute to standards and best practices for infrastructure, deployment, and operations.
• Share knowledge and enhance the operational maturity of the organization.
• A minimum of 4 years of experience in building and managing production infrastructure within cloud environments.
• Practical experience with AWS core services, including EC2, ECS/EKS, VPC, IAM, RDS, ElastiCache, ALB/NLB, and CloudWatch.
• Knowledge of containerized workloads and orchestration principles.
• Experience in supporting systems designed for high availability, fault tolerance, and controlled failure.
• Hands-on expertise with infrastructure as code tools, such as Terraform or CloudFormation.
• Experience in contributing to infrastructure migrations and production modifications.
• Proficient in troubleshooting production incidents related to networking, scaling, or service degradation.
• Experience managing infrastructure for financial systems or other high-reliability sectors is preferred.
• Background in managing infrastructure during growth phases or increasing complexity is preferred.
• Familiarity with observability and strong operational practices is preferred.
• Experience in simplifying, retiring, or rearchitecting existing infrastructure is preferred.
• Comfort in collaborating closely with a Rails-based application stack while maintaining tool-agnosticity is preferred.
• Generous PTO and company holiday policy.
• Company-paid Short Term Disability.
• 100% employer-covered health and dental insurance for direct employees.
• Higher-tier healthcare coverage available at the employee’s additional cost.
• Child dependent coverage available at the employee’s cost.
• Vision plan available at the employee’s additional cost.
• Child Care Benefits.
• Generous parental leave.
Helpware
Capgemini
Bitso
Peraton
Get handpicked remote jobs straight to your inbox weekly.