
Cloud Infrastructure Operations Engineer
Posted 22 hours ago

Posted 22 hours ago
This is a fully remote position, open to applicants in United States.
• Collaborate with the Cloud service team, internal departments, and cloud vendors to facilitate infrastructure delivery and deployment tasks.
• Execute operating system reinstalls and reboots for cloud servers and bare metal environments.
• Diagnose infrastructure challenges related to Linux systems, cloud servers, networking, and hardware anomalies.
• Manage operational data across internal systems, including fault ticketing systems, repair logs, and infrastructure tracking tools.
• Work alongside cloud providers to address infrastructure incidents and large-scale operational challenges.
• Engage in on-call rotations and assist with incident management activities.
• Oversee the health of cloud infrastructure, operational alerts, and asset utilization.
• Create or improve operational tools, scripts, and automation workflows using Shell or Python.
• Contribute to initiatives aimed at optimizing and standardizing operational processes.
• Conduct troubleshooting for server and cloud network issues, including TCP/IP, VLAN, DNS, and IPv6.
• Aid in the management of cloud fleet operations and infrastructure performance.
• Support the development of documentation, knowledge sharing, and operational reporting.
• Bachelor’s Degree in Computer Science, Electrical Engineering, or a related field.
• Experience in cloud infrastructure operations, server management, or data center environments.
• Strong troubleshooting and analytical capabilities in Linux and infrastructure settings.
• Familiarity with automated provisioning and OS deployment technologies, such as PXE, iPXE, or provisioning pipelines.
• Knowledge of public cloud platforms like Oracle, Amazon Web Services, Google, or Microsoft.
• Proficiency in understanding, executing, and writing Shell/Bash or Python scripts.
• Experience with automation and infrastructure management tools such as Ansible, GitLab CI/CD, Terraform, or cloud SDKs.
• Understanding of Linux systems, cloud networking, and server lifecycle management.
• Strong grasp of TCP/IP networking principles, including subnetting, VLANs, DNS, IPv6, and basic routing.
• Familiarity with infrastructure monitoring, operational tools, and incident management practices.
• Excellent communication, collaboration, and documentation skills.
• Experience with large-scale cloud fleet operations is preferred.
• Knowledge of GPU infrastructure support, firmware lifecycle management, or RDMA networking is a plus.
• Comprehensive health and wellness plans.
• Flexible working hours and remote work options.
• Opportunities for professional development and career advancement.
• Collaborative and inclusive work environment.
• Access to cutting-edge technology and resources.
Cint
Cint
Atmosera
Cint
Get handpicked remote jobs straight to your inbox weekly.