
Engineering Manager – AI Product
Posted Aug 4

Posted Aug 4
This is a fully remote position, open to applicants in Romania.
• Lead, mentor, and develop a team of 3 engineers dedicated to AI platform and product engineering.
• Drive the architectural vision for the AI product, focusing on inference performance, reliability, cost efficiency, and enhancing developer experience.
• Actively contribute to the design, implementation, and optimization of model serving, inference pipelines, retrieval, orchestration, and evaluation systems.
• Oversee the AI/ML platform comprehensively, including GPU capacity management, scheduling, model deployment and versioning, data and embedding pipelines, monitoring, and quality evaluation.
• Work collaboratively with Product, Infrastructure, Security, and Backend teams to set priorities and ensure seamless integrations.
• Conduct regular one-on-one meetings, performance evaluations, and provide technical coaching.
• Promote ownership, engineering excellence, and a culture of continuous improvement.
• Convert strategic objectives into actionable plans and ensure timely project delivery.
• Drive the evolution of the AI platform through informed technical decisions.
• Advocate for best practices in software development, automation, evaluation, and operational excellence.
• Over 7 years of experience in software engineering.
• At least 2 years in a technical leadership or engineering management capacity.
• Demonstrated experience in building and managing AI/ML platforms or inference infrastructure in production environments.
• Proficient in model serving, GPU workloads, latency optimization, and cost management.
• Solid experience in developing distributed backend systems.
• Strong proficiency in Python programming.
• Familiarity with systems programming languages such as Go or Rust is advantageous.
• Hands-on experience with model APIs, open-weight models, vLLM, TGI, Triton, retrieval/vector databases, and prompt or context engineering.
• Knowledgeable in AI product quality, evaluation frameworks, benchmarking, regression testing, and observability for non-deterministic systems.
• Significant expertise in cloud platforms, particularly AWS services including EC2, ECS, Lambda, IAM, and CloudWatch.
• Experience with GPU computing.
• Familiar with Kubernetes, containerization, and Infrastructure-as-Code tools such as Terraform, Pulumi, or AWS CDK.
• Understanding of CI/CD automation using GitHub Actions, CircleCI, or similar tools.
• Strong knowledge of Linux systems, networking fundamentals, and best practices in monitoring and alerting.
• Exceptional communication and leadership abilities.
• Strong technical judgment with the capability to balance coding responsibilities with leadership and strategic tasks.
• Proficiency in the English language is required.
• Must be able to work in alignment with the European time zone.
• Fully remote position.
• Flexibility with a remote-first approach.
• Opportunity to enhance expertise in LLMs, inference infrastructure, GPU computing, and large-scale distributed systems.
• Leadership, mentoring, and opportunities for strategic technical growth.
• Cross-functional collaboration with Product, Infrastructure, Security, and Backend teams.
Remote People
Bestow
Virta Health
Space Inch
Get handpicked remote jobs straight to your inbox weekly.