Remotery

Site Reliability Engineer – Guardicore AI Platform

Posted Jul 23

This is a fully remote position, open to applicants in Spain.

📋 Description

• Managing secure and highly available Kubernetes infrastructure for essential microservices, data pipelines, observability, and internal tools.

• Improving platform reliability, observability, security, performance, and cost-effectiveness.

• Offering support and advice to engineers and developers to boost their confidence in service performance.

• Leading intricate production investigations and promoting long-term enhancements.

• Utilizing LLMs and AI-driven automation to automatically resolve incidents and optimize operations.

• Collaborating with DevOps, Software, Data, AI, and Security engineering teams to analyze and troubleshoot complex issues.

• Engaging in on-call rotations, directing the restoration and repair of service-affecting problems.


⛳️ Requirements

• Over 3 years of experience in SRE, DevOps, or Platform Engineering, with a demonstrated history of mastering and diagnosing complex system architectures.

• Proven capability to design and execute a comprehensive monitoring and observability strategy utilizing tools such as Prometheus and Grafana.

• Practical experience with Kubernetes, Docker, Helm, and third-party cloud services (GCP, Azure, Linode, AWS) on Linux-based systems.

• Exceptional troubleshooting and problem-solving abilities across network, system, application, and database levels.

• Familiarity with GitOps, CI/CD, and Infrastructure as Code methodologies.

• Proficient in scripting and programming languages including Python, Go, and Bash.

• Employing AI tools in daily operational responsibilities and proactively suggesting initiatives to enhance platform automation.

• Exhibiting technical leadership and ownership in spearheading cross-team projects, defining tools, and constructing foundational frameworks.


🏝️ Benefits

• We prioritize your health, well-being, financial security, and life outside of work.

• FlexBase adapts to the requirements of your job.

• Our focus is on empowering employees to excel in their roles. We trust our exceptional team members to work in ways that best suit their needs: whether at home, in the office, or a blend of both.

People also viewed

CWILL21 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3722 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT22 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group22 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo23 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch23 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers