Senior Site Reliability Engineer

Posted 4 days ago

This is a fully remote position, open to applicants in Poland.

📋 Description

• Design solutions that enhance automation and efficiency for systems and teams.

• Streamline workflows, infrastructure, and applications.

• Work collaboratively on deployment, monitoring, and incident resolution.

• Enhance system monitoring to facilitate quicker error detection and resolution.

• Improve the performance and dependability of the virtualization platform.

• Create and maintain automated tools and scripts to ensure system reliability, deployment, and incident response.

• Engage in on-call rotations and lead the restoration and repair of service-impacting issues.

• Develop automation and tooling to minimize operational toil and enhance deployment safety.

• Contribute to capacity planning, autoscaling settings, and workload scheduling for AI compute infrastructure.

• Provide support and mentorship to fellow engineers.

• Foster continuous improvement and operational excellence throughout systems.


⛳️ Requirements

• Extensive experience in a SysAdmin (Linux/Unix Administration), DevOps, or SRE capacity.

• Proven experience with large-scale distributed systems.

• Proficient in Kubernetes and large-scale containerization technologies.

• Skilled in at least one programming language: Python or Golang.

• Experience with configuration management tools such as Terraform, SaltStack, or Ansible.

• Familiarity with defining SLOs.

• Knowledge of observability tools, including Prometheus and Grafana.

• Experience with distributed tracing.

• Ability to architect software and infrastructure at scale.

• Competence in developing automation and monitoring solutions.

• Strong collaboration skills with engineering teams that may not be familiar with SRE practices.


🏝️ Benefits

• Support for career advancement through GROW and Mentoring initiatives.

• Participation in internal development events such as the APEX Expo.

• Access to LinkedIn Learning resources.

• Opportunities to acquire new skills, explore various roles, and pursue diverse opportunities.

• 15-minute exploratory call with the Recruiter.

People also viewed

In All Media17 hours ago

DevOps Engineer – Cloud

BR flagBrazil, +5 more countriesFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity Group18 hours ago

SRE Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Fingerprint20 hours ago

Senior Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$152k – $205k/year
ApplyView job
Endava22 hours ago

Senior DevOps Engineer, Terraform

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
CVS Health23 hours ago

Staff DevSecOps Engineer, Health

US flagConnecticut, +3 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$130.3k – $260.6k/year
ApplyView job
GoFasti23 hours ago

Senior DevOps Engineer

Latin AmericaFull-timeDevOps & Site Reliability Engineer (SRE)$5,000 – $6,000/month
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers