
Senior Platform Engineer
Posted 4 days ago

Posted 4 days ago
This is a fully remote position, open to applicants in United States.
β’ Design, develop, and automate the Software Development Lifecycle and the infrastructure that supports Fieldwire.
β’ Take ownership of the reliability and performance of production systems, including incident response, root cause analysis, and proactive measures.
β’ Create and maintain systems for monitoring, alerting, and observability.
β’ Lead the troubleshooting and resolution of intricate application performance, scalability, and security challenges.
β’ Establish technical direction and make architectural decisions for platform infrastructure in collaboration with engineering leadership.
β’ Collaborate and influence engineering teams to construct a scalable platform.
β’ Mentor engineers on best practices for infrastructure, operational excellence, and secure system design.
β’ Propel continuous improvement of tools, processes, and standards.
β’ Participate in and assist in leading an on-call rotation, managing change and incident management processes.
β’ Spearhead Infrastructure as Code (IaC) projects and evaluate infrastructure designs.
β’ Design, develop, and support tools and environments for internal teams.
β’ Operate and enhance production systems, diagnosing and resolving live incidents.
β’ Provide support and guidance on internal development tools and workflows.
β’ Bachelor's or Master's degree in Computer Science or an equivalent level of professional experience.
β’ A minimum of 5-7 years of experience as a Platform Engineer, DevOps Engineer, or Site Reliability Engineer (SRE), with direct responsibility for production systems.
β’ Proficiency in at least one programming language, such as Golang, Java, Python, Ruby, or Rust.
β’ Demonstrated experience managing AWS cloud infrastructure and related services concerning CI/CD at scale.
β’ Extensive experience with Docker and containerization in production settings.
β’ Significant background in participating in and leading an on-call rotation, responding to and resolving production incidents.
β’ Experience in developing, refining, and promoting processes related to change and incident management.
β’ Strong expertise in Infrastructure-as-Code (IaC), such as AWS CDK or Terraform, with a focus on designing reusable and secure IaC patterns.
β’ Proven ability to mentor fellow engineers and influence technical decisions across teams.
β’ Familiarity with Rust and its ecosystem.
β’ Experience in designing and operating CI/CD systems that are integrated across multiple accounts and regions.
β’ Experience with package management systems and strategies.
β’ A history of leading cross-team infrastructure or reliability initiatives.
β’ A passion for tackling unique and challenging problems.
β’ Eligibility for a corporate bonus of up to 30%.
β’ Fully remote work opportunities within the United States.
Resend
Resend
Improvix Technologies
Improvix Technologies
Get handpicked remote jobs straight to your inbox weekly.