Remotery

Engineer II, Site Reliability

Posted 6 days ago

This is a fully remote position, open to applicants in United Kingdom.

📋 Description

• Design and implement automation and tools through software for essential solutions and services that support extensive distributed systems.

• Manage and engineer Linux systems across thousands of bare-metal servers and virtual environments.

• Take ownership of platform availability, latency, throughput, monitoring, incident response, deployment, and capacity planning.

• Engage in an on-call rotation.

• Diagnose server hardware issues.

• Ensure the platform operates reliably around the clock.

• Learn and advocate for new technologies and methodologies within the team.

• Acquire extensive exposure to the overall architecture and process workflow.

• Deliver minor development projects and occasionally larger initiatives.

• Utilize monitoring and telemetry tools such as ELK, Prometheus, Grafana, and Zabbix.

• Collect and evaluate operating system and application metrics for performance tuning and fault detection.

• Lead incident analysis, promote incident-response practices, link incidents to systemic issues, and drive resolutions.

• Collaborate with Site Reliability Engineers (SREs) and engineers distributed globally.

• Communicate and present the conventions followed by the reliability team.

• Leverage AI technologies to improve decision-making, streamline workflows and processes, enhance efficiency, and drive business outcomes.


⛳️ Requirements

• Bachelor's degree and/or equivalent experience in Computer Science.

• At least five years of experience in a large-scale production environment.

• Minimum of two years of experience in software engineering.

• At least two years of experience in one or more programming languages: C++, Java, Python, or Go.

• Familiarity with storage technologies such as SAN, NAS, NFS, Object Storage, FreeNAS, and iSCSI.

• Knowledge of infrastructure technologies including Linux, Windows, VMware, Docker, and Kubernetes.

• Experience in writing technical documentation.

• Configuration management experience with tools like Puppet, Chef, Ansible, or similar.

• Strong understanding of application design and operational trade-offs.

• Analytical abilities combined with a strong sense of urgency, ownership, and motivation.

• Capacity to work effectively in a diverse, team-oriented environment with SREs and engineers.

• Ability to communicate broadly and present recommended conventions.

• Proven experience using AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency, and drive business outcomes.


🏝️ Benefits

• Industry-leading compensation and equity awards.

• Comprehensive wellness programs for physical and mental health.

• Competitive vacation and holiday policies for relaxation.

• Paid parental and adoption leave.

• Professional development opportunities available to all employees, regardless of level or role.

• Employee Networks, geographic neighborhood groups, and volunteer opportunities to foster connections.

• Vibrant office culture featuring world-class amenities.

• Certified as a Great Place to Work™ worldwide.

People also viewed

CWILL15 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3716 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT17 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group17 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo17 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch18 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers