
Specialist II, Cloud Platform – Operations
Posted 5 days ago

Posted 5 days ago
This is a fully remote position, open to applicants in India.
• Design, develop, and sustain scalable cloud infrastructure on AWS and/or GCP.
• Create and manage Terraform Infrastructure-as-Code for reproducible and version-controlled deployments.
• Implement and oversee IAM solutions, including integrations with Okta and Auth0.
• Spearhead AIOps initiatives to automate monitoring, incident response, and remediation processes.
• Provide mentorship to junior engineers in cloud infrastructure, DevOps, and automation practices.
• Collaborate with development and operations teams to ensure reliability, security, and optimal performance.
• Establish CI/CD pipelines and automate deployment processes.
• Monitor, troubleshoot, and enhance the performance, costs, and security of cloud infrastructure.
• Participate in on-call rotations and manage incident response.
• 6-9 years of experience in DevOps, cloud infrastructure, or site reliability engineering.
• Strong practical experience with AWS and/or GCP cloud platforms, including multi-cloud and vendor-agnostic architecture patterns.
• Expert-level skills in Terraform for provisioning and managing infrastructure.
• Experience in implementing and managing identity and access management solutions, such as Okta and/or Auth0.
• Solid understanding of DevOps principles, including containerization, orchestration, and automation.
• Familiarity with AIOps concepts and tools for operational intelligence, predictive analytics, and automated remediation.
• Proficiency in Python, Bash, Go, or other similar scripting languages.
• Comprehensive knowledge of Docker and Kubernetes.
• Relevant cloud certifications (AWS, GCP, or similar).
• Experience with machine learning operations (MLOps) and the deployment of AI/ML infrastructure.
• Familiarity with monitoring tools like Dynatrace, New Relic, Prometheus, or the ELK stack.
• Understanding of security best practices, infrastructure security hardening, and vulnerability management.
• Experience in implementing incident management and disaster recovery strategies.
• Contributions to open-source infrastructure or DevOps projects.
• Equal opportunity employer.
• Inclusive environment for all employees.
Cube Care Company
ALB Conciergerie
ALB Conciergerie
Get handpicked remote jobs straight to your inbox weekly.