Remotery

Site Reliability Engineer – First-Level Operations & Support

Posted 2 days ago

This is a fully remote position, open to applicants in Ireland.

📋 Description

• Oversee production systems, applications, and infrastructure through the use of monitoring and alerting tools.

• Address alerts, investigate incidents, and implement corrective actions in accordance with established protocols.

• Conduct regular health checks and execute scheduled maintenance tasks.

• Analyze application, system, and server logs to diagnose issues.

• Restart services and perform operational tasks by utilizing documented runbooks.

• Manage and resolve support tickets while adhering to agreed service levels (SLAs).

• Escalate complex issues to second-level support or engineering teams when necessary.

• Track incidents until resolution and ensure accurate documentation is maintained.

• Perform daily operational checks to guarantee system stability and availability.

• Update and maintain operational documentation, runbooks, and troubleshooting guides.

• Participate in shift rotations, including weekends and public holidays.

• Assist in incident management activities and post-incident evaluations.

• Identify recurring issues and propose enhancements to improve reliability and efficiency.


⛳️ Requirements

• Minimum of 5 years of experience in Technical Support, Operations, NOC, SOC, or Site Reliability Engineering roles.

• Proficient troubleshooting, analytical, and problem-solving abilities.

• Familiarity with monitoring and alerting tools such as Zabbix, Grafana, and Prometheus.

• Experience in reviewing and analyzing application, system, and server logs.

• Knowledge of incident management and escalation procedures.

• Proficient in using ticketing systems like Jira.

• Experience with cloud platforms, especially Google Cloud.

• Understanding of Linux and common command-line utilities.

• Ability to adhere to structured procedures and operational runbooks.

• Strong attention to detail and a commitment to service reliability.

• Effective written and verbal communication skills.

• Capability to work independently and efficiently within a shift-based team.

• Availability for shifts, including weekends and public holidays.

• Ability to work within CET ± 2 hours.

• B2B contract outside IR35.


🏝️ Benefits

• Flexible working arrangements.

• Opportunity for repeat engagements based on performance.

• Access to CX guidance and market insights through our professional network.

• 6-month contract duration with automatic renewal.

• Clearly defined scope with no ambiguity regarding deliverables.

People also viewed

CWILL14 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3715 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT16 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group16 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo16 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch16 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers