
Site Reliability Engineer (SRE) – UI/UX
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in Canada.
• Assist in the deployment, operation, and ongoing maintenance of production services running on Kubernetes.
• Monitor the health, availability, and performance of services.
• Investigate and resolve production incidents by utilizing logs, monitoring, and debugging tools.
• Conduct log analysis and incident debugging with Splunk.
• Identify service-related issues and work with engineering teams to facilitate timely resolutions.
• Engage in incident response and activities related to production support.
• Perform initial debugging of UI-related issues that involve Web Components.
• Contribute to service reliability and initiatives for continuous improvement.
• Provide assistance with CI/CD pipelines and operations of cloud-native applications as needed.
• Operate effectively within a client-directed backlog and adhere to established priorities.
• A minimum of 4 years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Support, or a similar role.
• Practical experience in supporting the deployment, operations, and ongoing maintenance of production services on Kubernetes.
• Background in monitoring service health, troubleshooting production issues, and ensuring service reliability.
• Proficiency in using Splunk for log analysis and debugging incidents.
• Experience in participating in production incident response and conducting root-cause analysis.
• Familiarity with Web Components and the ability to perform initial debugging of UI-related issues.
• Strong troubleshooting, analytical, and problem-solving capabilities.
• Experience collaborating with software engineering and cross-functional teams.
• Ability to work independently and efficiently within a client-directed backlog.
• Excellent proficiency in written and spoken English, at least at a B2 level.
• Preferred: experience with supporting CI/CD pipelines.
• Preferred: knowledge of multi-tenant services.
• Preferred: experience in cloud-native application operations.
• Preferred: background in supporting high-availability enterprise or SaaS platforms.
• Preferred: familiarity with additional monitoring and observability tools.
• Preferred: experience with cloud platforms like AWS, Azure, or GCP.
• Preferred: knowledge of Docker and Helm.
• Competitive salary.
• Laptop provided.
• Opportunities for professional development and training.
• Work with advanced cloud and container technologies.
• Flexible work arrangements available.
• Collaborative team environment.
• Opportunity to impact organization-wide digital transformation initiatives.
CVS Health
Devoteam
Aspirion
Goodgame Studios
Get handpicked remote jobs straight to your inbox weekly.