
Member of Engineering – Compute
Posted Jul 18

Posted Jul 18
This is a fully remote position, open to applicants in Europe.
• Create and implement an internal scheduling system aimed at optimizing GPU utilization.
• Develop APIs and tools that facilitate the management of GPU workload lifecycles and assist in troubleshooting failures.
• Design and enhance the inference control plane to accelerate model deployment and improve the efficiency of inference request handling.
• Collaborate with research teams to continually enhance research velocity.
• Proficient programming abilities in Go or comparable languages.
• Solid background in systems engineering, including distributed systems, schedulers, control planes, or high-throughput data planes.
• Hands-on experience with Kubernetes internals, such as controllers, informers, and operators—not just deploying applications on it.
• A focus on observability and debuggability, ensuring the system is easy to navigate during the debugging of production issues.
• Preferred: experience in systems that handle large-scale inference requests.
• Fully remote work with flexible working hours.
• 37 days of vacation and holidays per year.
• Health insurance allowance for you and your dependents.
• 16 weeks of flexible, fully-paid parental leave.
• Allowances for well-being, continuous learning, and home office setup.
• Equipment provided by the company.
• Regular team gatherings.
• A diverse and inclusive, people-first culture.
Get handpicked remote jobs straight to your inbox weekly.