
Staff Software Engineer
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in United States.
• Spearhead the transformation of DevOps through the design and enhancement of CI/CD pipelines, strategies for Kubernetes container orchestration, and robust cloud infrastructure patterns.
• Design and implement secure infrastructure with a focus on secrets management, least-privilege IAM, vulnerability scanning, and compliance protocols.
• Utilize and critically assess AI-assisted development tools for platform tasks and ensure responsible AI practices.
• Take ownership of platform reliability by establishing SLOs/SLAs, guiding incident response efforts, and minimizing MTTD/MTTR through effective observability and alerting mechanisms.
• Create scalable infrastructure-as-code modules using Terraform/Pulumi and address overarching technical debt.
• Define the platform architecture across various teams and domains, ensuring alignment with engineering and organizational strategies.
• Deliver impactful solutions by leveraging emerging platform technologies.
• Enhance cloud architecture to improve reliability, cost-effectiveness, and performance.
• Oversee cross-functional dependencies, risks, and change management processes.
• Provide mentorship to senior engineers via code reviews, architectural guidance, documentation, recruitment, and team development.
• 8-12+ years of experience in Platform Engineering, SRE, or DevOps roles.
• Extensive knowledge of Infrastructure as Code utilizing Terraform, Pulumi, or CloudFormation.
• Proficient in container orchestration with Kubernetes and Docker.
• Experience in designing and optimizing CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or similar tools.
• Expertise in cloud platforms such as AWS, GCP, or Azure, focusing on architecture, cost optimization, and IAM practices.
• Familiarity with observability tools like Sumo, Datadog, Grafana, Prometheus, or similar platforms.
• Proven track record of enhancing reliability metrics, including MTTD/MTTR and uptime.
• Knowledge of security-focused infrastructure design that incorporates secrets management, least privilege, vulnerability scanning, and compliance awareness.
• Strong capability in utilizing AI tools for code development, triage, and maintenance, with the ability to critically assess AI-generated results.
• Proficient scripting and automation skills in Python, Go, or Bash.
• Experience leading cross-functional technical initiatives and fostering organizational adoption.
• Understanding of AI/ML infrastructure considerations.
• Experience managing production reliability for a critical business service.
• Excellent debugging skills and the ability to resolve complex issues.
• Industry-competitive pay.
• Equity in the company.
• 401(k) plan with a 4% immediate vesting match.
• 16-18 weeks of paid parental leave for birthing parents.
• 10 weeks of paid parental leave for non-birthing parents.
• Coverage for infertility treatments.
• Comprehensive medical, dental, and vision plans.
• 100% employer-paid dental and vision for individual coverage.
• Affordable medical options.
• 10 complimentary mental health visits.
• Pet insurance available.
• Generous paid time off (PTO).
• Return to Office (RTO) policy for full-time exempt employees.
• Sick leave provisions.
• Paid holidays.
• Personal holiday.
• Annual Evolve travel credit after one year of service.
• Discounts on stays at all Evolve properties.
• Exceptional onboarding programs.
• Access to learning and development resources.
• Participation in Employee Resource Groups.
Cloudera
Stellar Cyber
Pragmatike
Pragmatike
Get handpicked remote jobs straight to your inbox weekly.