
Junior/Senior Site Reliability Engineer – Night Shift
Posted 20 hours ago

Posted 20 hours ago
This is a fully remote position, open to applicants in India.
• Design, implement, and manage scalable and highly available systems on Azure Cloud.
• Monitor system performance, troubleshoot issues, and ensure uptime and reliability.
• Manage and optimize Kubernetes clusters and containerized workloads using Docker.
• Build and maintain robust CI/CD pipelines utilizing GitHub Actions and related tools.
• Implement infrastructure as code and automate deployments through Helm Charts.
• Work with distributed systems such as Kafka, Redis, PostgreSQL, and Hadoop/HDFS.
• Configure and manage Cloudflare for enhanced performance, security, and traffic routing.
• Set up monitoring, alerting, and observability with tools like Grafana.
• Collaborate with development teams to enhance system reliability and deployment practices.
• Conduct root cause analysis (RCA) and implement preventive measures.
• Ensure adherence to security best practices and compliance across the infrastructure.
• 3–9 years of experience in SRE/DevOps or related fields.
• Strong hands-on experience with Azure Cloud services.
• Solid background in Linux system administration.
• Expertise in Docker and Kubernetes, including deployment, scaling, and troubleshooting.
• Experience with Kafka, Redis, and PostgreSQL.
• Working knowledge of the Hadoop ecosystem, including HDFS and Hadoop.
• Familiarity with Cloudflare for CDN, security, and DNS management.
• Proficiency in CI/CD tools such as GitHub and GitHub Actions.
• Experience with Helm Charts and Kubernetes deployments.
• Strong understanding of monitoring and logging tools, such as Grafana.
• Full-stack benefits for health, wealth, and wellbeing to keep you thriving.
The Codest
IRIUM
Sólides
Verity Group
Get handpicked remote jobs straight to your inbox weekly.