Remotery

Site Reliability Engineer

Posted 3 hours ago

This is a fully remote position, open to applicants in India.

📋 Description

• Participate in an on-call rotation addressing production availability incidents and assisting service engineers with customer-related issues.

• Utilize on-call shifts to mitigate the recurrence of incidents.

• Manage infrastructure using Ansible, Puppet, Terraform, and Kubernetes.

• Set up monitoring and alerting systems to identify symptoms rather than full outages.

• Keep a detailed record of all actions taken, ensuring that insights become repeatable practices and automation.

• Enhance the deployment process.

• Design, construct, and sustain core infrastructure capable of scaling to hundreds of thousands of concurrent users.

• Troubleshoot production issues across various services and stack levels.

• Strategize infrastructure expansion.

• Code infrastructure automation utilizing Ansible and Terraform.

• Enhance Prometheus monitoring or create new metrics.

• Assist release managers in deploying and resolving issues with new application software versions.

• Plan and implement the migration from AWS virtual machines to cloud-native, containerized deployments on Kubernetes (EKS).

• Build relationships with product teams and establish their SRE KPIs.


⛳️ Requirements

• Embrace a cloud-first mindset, regardless of the public cloud provider.

• Prioritize security in all considerations.

• Have a comprehensive understanding of systems, including edge cases, failure modes, behaviors, and specific implementations.

• Proficient in both Linux and Windows operating systems.

• Familiar with configuration management systems such as Ansible or Puppet.

• Strong programming capabilities in Python, Java, Golang, or Node.js.

• Ability to collaborate and communicate in an asynchronous manner.

• Thoroughly document all work.

• Possess a proactive attitude and a willingness to resolve issues in dysfunctional systems.

• Experience with Nginx, HAProxy, Docker, Kubernetes, Terraform, or similar technologies.


🏝️ Benefits

• Competitive salary and performance-based bonuses.

• Comprehensive health, dental, and vision insurance.

• Flexible working hours and remote work options.

• Opportunities for professional development and continuous learning.

• Collaborative and inclusive work environment.

People also viewed

Endava2 hours ago

Senior DevOps Engineer, Dynatrace

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Jones Lang LaSalle Americas, Inc.2 hours ago

Reliability Engineer

US flagIllinois, +1 more stateFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $120k/year
ApplyView job
NVIDIA2 hours ago

Service Reliability Engineer

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$168k – $333.5k/year
ApplyView job
Entarian2 hours ago

DevSecOps Engineer – Mid

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
BeyondTrust2 hours ago

Senior DevOps Engineer

CA flagCanada OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ontrac Solutions3 hours ago

Site Reliability Engineer

PK flagPakistan OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers