
Senior Site Reliability Engineer – Compute Platform Services Team
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in India.
• Work collaboratively with support, operations, and engineering teams to investigate and resolve intricate issues.
• Create processes, plans, and infrastructure to deploy new software components and updates securely and efficiently at scale.
• Participate in on-call rotations, providing guidance for the restoration and repair of service-affecting problems.
• Enhance system monitoring and analysis platforms to speed up error detection and resolution.
• Boost automation, operational excellence, and support for applications and infrastructure that interact with customers.
• Over 7 years of relevant experience.
• Bachelor's degree in Computer Science or a related discipline.
• Advanced expertise in Systems Engineering, DevOps, or Software Engineering.
• Experience with large-scale distributed systems.
• Capability to troubleshoot systemic issues and create large-scale automations.
• Proficiency in Python or Golang.
• Practical experience with SaltStack, Ansible, and Terraform.
• Expertise in observability or monitoring tools such as Prometheus, Grafana, ELK/OpenSearch, Datadog, and Splunk.
• Familiarity with cloud platforms such as AWS, GCP, Azure, or their equivalents.
• Proficiency in Linux systems administration, configuration management, performance optimization, and hardware engineering.
• Comprehensive health, well-being, financial, and life benefits.
• FlexBase flexible work arrangements: from home, in an office, or a blend of both.
CVS Health
Devoteam
Aspirion
Goodgame Studios
Get handpicked remote jobs straight to your inbox weekly.