
Senior DevOps Engineer, Infrastructure – Reliability
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in Florida.
• Develop scalable Infrastructure-as-Code frameworks utilizing Terraform.
• Take ownership of and advance the Kubernetes platform, which may involve EKS or self-managed Kubernetes setups.
• Enhance CI/CD pipelines to boost deployment frequency, lead time, and confidence in releases.
• Design and implement secure networking, IAM, and secrets management across various environments.
• Enhance metrics, logging, and tracing capabilities using DataDog.
• Reduce cloud expenses through rightsizing, autoscaling, and architectural enhancements.
• Establish disaster recovery, backup, and multi-region resilience strategies.
• Transform brittle or manual infrastructure into automated, testable, and reproducible systems.
• Introduce infrastructure tools and architectural modifications through documentation, workshops, and direct support.
• Collaborate with engineering teams to minimize CI/CD, deployment, and cloud-related challenges.
• Convey technical trade-offs to engineering and product stakeholders.
• Work with technologies including AWS, Kubernetes, ArgoCD, Terraform, GitHub Actions, DataDog, PostgreSQL, Kafka, Redis, Bash, Python, TypeScript, and JavaScript.
• 8+ years of experience in DevOps, SRE, or infrastructure engineering.
• Demonstrated experience in designing and managing production Kubernetes environments at scale.
• Extensive hands-on knowledge of AWS infrastructure and cloud networking.
• Strong background in creating and maintaining Terraform modules within large cloud environments.
• Proven ownership of CI/CD systems with measurable enhancements in DORA metrics.
• Experience leading incident response efforts and achieving significant postmortem results.
• Solid understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).
• Proven capability to modernize legacy infrastructure and reduce manual operational burdens.
• Track record of successfully managing scoped infrastructure projects from unclear beginnings to production without requiring daily oversight.
• Ability to foster trust among teams while improving reliability standards.
• Preferred qualifications include application development, high-throughput Kafka, PostgreSQL/Redis optimization, autoscaling, service mesh, internal developer platforms, zero-trust networking, policy-as-code, multi-region systems, and reliability frameworks.
• Health Care Plan (Medical, Dental & Vision)
• Retirement Plan (401k)
• Life Insurance
• Flexible Paid Time Off
• 9 paid Holidays
• Family Leave
• Remote work options
• Hybrid work arrangement for Orlando Associates
• Complimentary Food & Snacks (Orlando)
• Wellness Resources
knowmad mood
RealTime eClinical Solutions
Koniag Government Services
Koniag Government Services
Get handpicked remote jobs straight to your inbox weekly.