
Software Engineer
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in United Kingdom.
• Develop and sustain automation processes that ensure the reliability of a large VR application in production, encompassing release pipelines, build health, crash analysis, and incident detection.
• Enhance and maintain an AI-driven code repair system that autonomously creates and implements fix diffs.
• Create tools to automatically identify broken builds, determine root-cause changes, and suggest or execute fixes.
• Monitor production quality metrics and address any regressions and outages promptly.
• Alleviate the manual on-call workload through automation, aiming to reduce recurring operational tasks by 80–90%.
• Execute necessary infrastructure migrations to ensure the functionality of CI/CD pipelines as upstream dependencies become obsolete.
• Collaborate with the runtime and KTLO teams.
• Design and manage systems for a substantial UGC application across VR headsets, mobile devices, and PC via cloud streaming.
• Over 8 years of professional experience in software engineering, or an equivalent background.
• Demonstrated expertise in building and managing CI/CD, build, release, and cloud deployment pipelines at scale.
• Proficient in operating cloud services and server-side fleets in a production environment, focusing on reliability, capacity, and latency.
• Experience in developing or managing AI-assisted developer tools or agents that generate or repair code.
• Skilled in creating tools that detect broken builds and trace failures to their root causes.
• Experienced in production monitoring, crash analysis, and incident response for large-scale, multi-surface applications.
• A proven history of minimizing operational and on-call burdens through automation.
• Experience in executing infrastructure or dependency migrations without disrupting downstream CI/CD.
• Knowledge in cloud gaming or application streaming, or remote rendering.
• Familiarity with asset delivery or CDN pipelines at scale.
• Expertise in capacity, latency, or session-orchestration monitoring for streamed workloads.
• Background in operating live-service or large-scale production applications (Live Ops).
• Acquainted with large monorepo build systems and dependency management.
• Experience in designing self-healing or auto-remediation systems.
• Competitive salary.
• Healthcare contribution.
• Inclusion in the company pension scheme.
• Provision of work laptop and phone.
• 25 days of annual leave (pro-rata).
• Paid bank holidays.
• Opportunities for career advancement for top performers.
Wellspring Worldwide
Iambic Therapeutics
Quality Digital
Get handpicked remote jobs straight to your inbox weekly.