
Senior Cloud Database Engineer
Posted Jul 27

Posted Jul 27
This is a fully remote position, open to applicants in India.
• Take charge of the daily operations, health, availability, and ongoing enhancement of Toku's production MySQL databases hosted on AWS RDS, ensuring consistent reliability across our global SaaS platform.
• Assume responsibility for production incidents, executing troubleshooting and root cause analysis, swiftly restoring service, and implementing permanent improvements to minimize recurring operational challenges.
• Recognize repetitive operational tasks, assess where automation provides maximum value, and develop scalable workflows and tools that mitigate manual effort, operational risk, and engineering overhead.
• Enhance the management of database changes, production data corrections, and operational requests by introducing safer, more auditable, and scalable engineering processes.
• Collaborate closely with software engineers to comprehend application behavior, investigate production issues, analyze logs, and strengthen the interaction between engineering and database operations.
• Monitor and refine database performance through query optimization, indexing strategies, capacity planning, and proactive performance evaluation.
• Manage and continually enhance AWS database services, focusing primarily on Amazon RDS while supporting Toku's transition towards Amazon Aurora.
• Utilize monitoring platforms such as Datadog and AWS Performance Insights to identify trends, troubleshoot production issues, and enhance operational visibility.
• Implement secure database practices, strengthen operational governance, support audit readiness, and assist in reducing operational risks associated with production database management.
• Support and consistently enhance high availability and disaster recovery capabilities while leveraging modern AWS managed database services.
• Question existing workflows, identify opportunities to reduce operational toil, and contribute ideas to improve the long-term scalability of the database platform.
• Participate in a shared on-call rotation (approximately two weeks per month), addressing production incidents outside regular business hours as needed.
• 5–8 years of hands-on experience managing production databases within AWS cloud environments.
• Strong practical experience with Amazon RDS.
• Extensive production experience in administering and optimizing MySQL databases.
• Experience managing cloud-based database services rather than traditional on-premise database infrastructure.
• Practical experience in automating operational tasks through scripting languages such as Python, SQL, or similar.
• Experience supporting mission-critical production systems, engaging in incident response, and operating within an on-call setting.
• Strong expertise in query optimization, indexing, performance troubleshooting, and root cause analysis.
• Good understanding of AWS networking essentials, including security groups and the infrastructure concepts necessary for troubleshooting production database environments.
• Experience collaborating closely with software engineering teams, with a solid grasp of modern software development lifecycles and the ability to work effectively across engineering and infrastructure.
• Proven ability to investigate unfamiliar production challenges, learn independently, and devise practical solutions without solely relying on predefined procedures.
• A strong sense of ownership, excellent communication skills, and a proactive approach to driving work to completion while keeping stakeholders informed.
• Curiosity and initiative to rapidly understand complex production environments, continually improve existing processes, and embrace new technologies as the platform evolves.
• Paid time off.
• Flexible working hours.
• Professional development opportunities.
Sigma Software Group
Plain Concepts
GitLab
Get handpicked remote jobs straight to your inbox weekly.