Senior DevOps Engineer, Infrastructure – Reliability

Posted 1 day ago

This is a fully remote position, open to applicants in Florida.

📋 Description

• Develop scalable Infrastructure-as-Code frameworks utilizing Terraform.

• Take ownership of and advance the Kubernetes platform, which may involve EKS or self-managed Kubernetes setups.

• Enhance CI/CD pipelines to boost deployment frequency, lead time, and confidence in releases.

• Design and implement secure networking, IAM, and secrets management across various environments.

• Enhance metrics, logging, and tracing capabilities using DataDog.

• Reduce cloud expenses through rightsizing, autoscaling, and architectural enhancements.

• Establish disaster recovery, backup, and multi-region resilience strategies.

• Transform brittle or manual infrastructure into automated, testable, and reproducible systems.

• Introduce infrastructure tools and architectural modifications through documentation, workshops, and direct support.

• Collaborate with engineering teams to minimize CI/CD, deployment, and cloud-related challenges.

• Convey technical trade-offs to engineering and product stakeholders.

• Work with technologies including AWS, Kubernetes, ArgoCD, Terraform, GitHub Actions, DataDog, PostgreSQL, Kafka, Redis, Bash, Python, TypeScript, and JavaScript.


⛳️ Requirements

• 8+ years of experience in DevOps, SRE, or infrastructure engineering.

• Demonstrated experience in designing and managing production Kubernetes environments at scale.

• Extensive hands-on knowledge of AWS infrastructure and cloud networking.

• Strong background in creating and maintaining Terraform modules within large cloud environments.

• Proven ownership of CI/CD systems with measurable enhancements in DORA metrics.

• Experience leading incident response efforts and achieving significant postmortem results.

• Solid understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).

• Proven capability to modernize legacy infrastructure and reduce manual operational burdens.

• Track record of successfully managing scoped infrastructure projects from unclear beginnings to production without requiring daily oversight.

• Ability to foster trust among teams while improving reliability standards.

• Preferred qualifications include application development, high-throughput Kafka, PostgreSQL/Redis optimization, autoscaling, service mesh, internal developer platforms, zero-trust networking, policy-as-code, multi-region systems, and reliability frameworks.


🏝️ Benefits

• Health Care Plan (Medical, Dental & Vision)

• Retirement Plan (401k)

• Life Insurance

• Flexible Paid Time Off

• 9 paid Holidays

• Family Leave

• Remote work options

• Hybrid work arrangement for Orlando Associates

• Complimentary Food & Snacks (Orlando)

• Wellness Resources

People also viewed

knowmad mood16 hours ago

Consultor/a DevSecOps – AWS

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
RealTime eClinical Solutions23 hours ago

Principal DevOps Architect

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$155k – $195k/year
ApplyView job
Koniag Government Services1 day ago

DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Koniag Government Services1 day ago

Senior AWS DevOps Engineer – AWS, Kubernetes, HCP, CI/CD, Observability, AI-focus

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ASRC Federal1 day ago

Senior DevOps Administrator – Supporting NASA

US flagCalifornia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Nelios1 day ago

DevOps Engineer, Cloud Infrastructure

GR flagGreece OnlyPart-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers