
Site Reliability Engineer
Posted Sep 9

Posted Sep 9
This is a fully remote position, open to applicants in United States.
• Ensure high availability of the site while delivering an exceptional customer experience
• Address intricate technical challenges
• Develop and enhance operational procedures
• Facilitate effective communication during incidents
• Engage in active incident management
• Create in-house tools and utilize AI-driven workflows
• Deliver frontline technical support for cloud infrastructure, databases, and network systems
• Analyze monitoring data and respond to patterns, outliers, and anomalies
• Collaborate with development teams to minimize issue detection and resolution times
• Oversee incident bridges with on-call support teams and executives during outages
• Automate reporting of incident cases through ticket tracking systems
• Conduct research, proofread, and write technical documentation
• Establish and maintain context for AI-driven workflows
• Assist in automation and tool development to alleviate repetitive tasks
• Contribute both independently and as part of a team in meetings and projects
• Work during some holidays as part of a 24x7x365 Command Center operation
• Over 4 years of experience in a continuous, high-availability enterprise production environment
• More than 4 years of experience managing and monitoring AWS technologies and services using tools like New Relic or AWS Application Signals
• At least 2 years of experience in an enterprise-level NOC or Command Center environment
• Over 2 years of experience in developing AI-driven workflows
• Familiarity with prompt engineering, context management, and AI SDLC best practices
• More than 4 years of experience in scripting, automation, and software development
• Proficiency in programming languages such as Python, Java, Bash, and PowerShell
• Ability to think logically through new technologies, systems, concepts, and procedures
• Competence in using reports and data to enhance operational outcomes
• Capability to document issues following established processes
• Understanding of TCP/IP LAN/WAN networking technologies and troubleshooting methods
• Awareness of virtualization technologies
• Knowledge of hardware- or software-based firewalls, load balancers, intrusion detection systems, and proxy servers
• Familiarity with relational and NoSQL database query languages, including MSSQL, MySQL, Cassandra, and CouchDB
• Knowledge of CI/CD tools such as Harness, Chef, or Puppet
• Willingness to cover additional shifts, including days, weekends, and holidays
• Availability to work the night shift from 10 pm–6 am MST, Monday through Thursday
• A background check is required for job offers
• Eligibility for bonuses
• Equity options
• Comprehensive health, dental, and vision insurance
• Flexible work location: nearest office, home, or hybrid (subject to restrictions)
• Competitive salary
• Collaborative work environment
• Opportunity to make a significant impact
FourEnergy GmbH
ICF
Mastercam
C&S Informática
Get handpicked remote jobs straight to your inbox weekly.