
DevOps, Platform Operations Engineer
Posted Sep 2

Posted Sep 2
This is a fully remote position, open to applicants in Philippines.
• Oversee the production and staging infrastructure of the SaaS platform, which includes cloud servers and services, databases, domains and DNS, SSL certificates, storage, networking, environment configuration, application services, user access and permissions, as well as infrastructure scaling.
• Continuously monitor system health to ensure that the infrastructure is properly configured as the SaaS platform expands.
• Deploy validated SaaS platform code reliably into both staging and production environments.
• Design and maintain automated CI/CD deployment pipelines.
• Ensure a clear distinction between Development, Staging, and Production environments, allowing for rapid rollback of releases when needed.
• Set up monitoring and alerting systems for application availability, server health, database and API performance, error rates, resource utilization, failed processes, integration failures, and security incidents.
• Serve as the primary technical contact for operational challenges.
• Troubleshoot issues, restore services, and document root causes effectively.
• Create Root Cause Analyses for major incidents and communicate prevention strategies.
• Manage backup and recovery systems, which include automated backups, geographically/reliably stored backups, documented recovery protocols, periodic testing, and rebuild capabilities.
• Keep a documented Disaster Recovery Procedure up to date.
• Uphold cloud and application security protocols, including access control, MFA, secrets management, API credentials, encryption, firewall/security configuration, vulnerability monitoring, security updates, logging, production access management, and backup security.
• Collaborate closely with product and development teams to ensure the platform remains secure, available, fast, and reliable.
• A minimum of 5 years of practical experience in DevOps, Platform Operations, Cloud Engineering, or a related field.
• Extensive experience with Microsoft Azure, covering infrastructure, monitoring, networking, security, and production deployments.
• Proficiency with Microsoft Entra ID, access controls, and identity management.
• Experience managing Linux servers, web applications, and production environments.
• Familiarity with Git/GitHub, CI/CD pipelines, Docker, and application deployments.
• Knowledge of DNS, SSL/TLS, SQL databases, backups, REST APIs, and third-party integrations.
• Experience in cloud monitoring, logging, incident response, troubleshooting, and root-cause analysis.
• Strong understanding of infrastructure and application security principles.
• Proficiency in scripting languages such as Python, Bash, PowerShell, or similar.
• Excellent problem-solving abilities with a methodical and hands-on approach.
• Ability to work independently, take ownership, and articulate technical issues in simple terms.
• Comfort in utilizing AI tools as part of daily responsibilities.
• Demonstrated enthusiasm, creativity, and genuine passion for the role and responsibilities.
• High level of engagement, initiative, and ownership in completing tasks and contributing to team goals.
• Proactive in recognizing opportunities, resolving issues, and offering ideas beyond assigned tasks.
• Brings a positive, collaborative, and results-oriented attitude to the workplace.
• Experience with Terraform or other Infrastructure as Code tools is advantageous.
• Comprehensive health insurance plans.
• Flexible working hours and remote work options.
• Opportunities for professional development and growth.
• Collaborative and innovative work environment.
• Competitive salary and performance-based bonuses.
Quvia
Anyone AI
Manulife
Accenture Federal Services
Get handpicked remote jobs straight to your inbox weekly.