Remotery

Senior Site Reliability Engineer – Guardicore AI Platform

Posted Jul 23

This is a fully remote position, open to applicants in Spain.

📋 Description

• Managing secure and highly available Kubernetes infrastructure for essential microservices, data pipelines, observability, and internal tools.

• Improving platform reliability, observability, security, performance, and cost-effectiveness.

• Offering support to engineers and developers to boost their confidence in service performance.

• Leading intricate production investigations and facilitating long-term enhancements.

• Utilizing LLMs and AI-driven automation to automatically resolve incidents and optimize operations.

• Collaborating with DevOps, Software, Data, AI, and Security engineering teams to analyze and resolve complex issues.

• Engaging in on-call rotations, overseeing the restoration and repair of service-impacting problems.


⛳️ Requirements

• Over 5 years of experience in SRE, DevOps, or Platform Engineering, showcasing a strong record of mastering and troubleshooting complex system architectures.

• Proven ability to design and implement a thorough monitoring and observability strategy using tools such as Prometheus and Grafana.

• Extensive hands-on experience with Kubernetes, Docker, Helm, and various cloud services (GCP, Azure, Linode, AWS) on Linux-based systems.

• Outstanding troubleshooting and problem-solving abilities across network, system, application, and database layers.

• Familiarity with GitOps, CI/CD, and Infrastructure as Code practices.

• Proficient in scripting and programming with Python, Go, and Bash.

• Apply AI tools in everyday operational tasks and proactively suggest initiatives to enhance platform automation.

• Exhibit technical leadership and ownership in driving cross-team initiatives, defining tools, and constructing foundational frameworks.


🏝️ Benefits

• We prioritize your health, well-being, financial stability, and life beyond the workplace. Check out our benefits.

• Akamai's FlexBase program exemplifies our commitment to offering employees an outstanding workplace experience. It's about empowering employees to perform their best work in a setting that suits them.

• We trust our exceptional employees to work in ways that align with their preferences: whether at home, in an office, or a mix of both.

People also viewed

CWILL21 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3722 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT22 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group22 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo23 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch23 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers