
Senior Site Reliability Engineer
Posted Jul 18

Posted Jul 18
This is a fully remote position, open to applicants in Argentina.
• Design, develop, and scale Kubernetes infrastructure for secure, multi-tenant, high-availability applications.
• Build and manage AI tooling infrastructure.
• Optimize and sustain CI/CD pipelines.
• Implement progressive delivery methodologies.
• Enhance Infrastructure as Code using Terraform, Helm, and Argo CD.
• Operate and refine streaming and analytics infrastructure.
• Incorporate automated testing within the CI/CD lifecycle.
• Enhance system observability.
• Lead incident response efforts and conduct postmortems.
• Over 6 years of experience in SRE, DevOps, or Infrastructure positions, with extensive production Kubernetes expertise.
• Practical experience in integrating AI/LLM tooling into engineering or operational processes.
• Demonstrated success in constructing CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI, or similar tools).
• Strong understanding of Kubernetes internals and managed services such as EKS, GKE, or AKS.
• Proficiency in Infrastructure as Code (Terraform, Helm, Pulumi) and GitOps methodologies.
• Skilled in Python, Bash, or Go programming languages.
• Familiarity with observability tools (Prometheus, Grafana, Datadog, OpenTelemetry).
• Experience with production-level systems like Kafka, Flink, and ClickHouse.
• Excellent communication and collaboration skills across teams.
• Competitive salary
• Stock options
• Health benefits
• Unlimited PTO
• Parental leave
• Tuition reimbursements
The Codest
IRIUM
Sólides
Resilinc
Get handpicked remote jobs straight to your inbox weekly.