Remotery

Senior Site Reliability Engineer – Cloud Platform

Posted Aug 18

This is a fully remote position, open to applicants in Philippines.

📋 Description

• Ensure the reliability, availability, and performance of both production and pre-production environments.

• Oversee platform health and enhance alerting, automation, and operational processes.

• Address production incidents, engage in root cause analysis, and implement lasting improvements.

• Design, construct, and refine observability solutions utilizing metrics, logs, traces, and dashboards.

• Collaborate with software engineers to enhance application reliability throughout the development lifecycle.

• Create and maintain operational documentation, troubleshooting guides, and runbooks.

• Automate repetitive operational tasks to boost efficiency and minimize manual intervention.

• Participate in on-call rotations while continuously refining incident response processes.

• Advocate for reliability engineering principles, operational excellence, and ongoing improvement across engineering teams.


⛳️ Requirements

• Bachelor's or Master's degree in Engineering, Computer Science, or a related discipline.

• Extensive experience in operating Kubernetes or other container orchestration platforms.

• Background in supporting large-scale production services.

• Practical experience with AWS.

• Familiarity with Prometheus, Grafana, and ELK.

• Strong scripting abilities in Bash, Python, or Go.

• Experience in managing Linux-based production environments.

• Knowledge of Infrastructure as Code or configuration management tools like Terraform or Ansible.

• A solid understanding of networking fundamentals, including TCP/IP, DNS, load balancing, and routing.

• Exceptional troubleshooting, communication, and collaboration skills.

• A proactive approach with a passion for automation and reliability.

• Nice to have: experience with SIP or VoIP technologies.

• Nice to have: familiarity with MySQL or PostgreSQL.

• Nice to have: experience with Redis or other NoSQL databases.


🏝️ Benefits

• Long-term, full-time collaboration.

• Flexible remote working environment.

• Opportunities for professional development, including training and technical learning.

• Chance to work with innovative cloud technologies utilized by customers globally.

• A collaborative engineering culture centered on knowledge sharing and continuous improvement.

• Provision of modern Apple equipment.

• An inclusive and respectful workplace regardless of gender, ethnicity, or background.

People also viewed

HubSpot11 hours ago

Principal Software Engineer, Developer Acceleration – Release Engineering

IE flagIreland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$242k – $387.2k/year
ApplyView job
InfluxData11 hours ago

DevOps Engineer

GB flagUnited Kingdom, +7 more countriesFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
LocalStack12 hours ago

Senior DevOps Engineer

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€50k – €78k/year
ApplyView job
Amigo Tech13 hours ago

DevOps Engineer, Mid-Level

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Penn Interactive13 hours ago

Senior Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$145k – $193k/year
ApplyView job
Smarthis13 hours ago

Senior Cloud Platform – DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers