
Engineering Manager – AI Product
Posted Aug 4

Posted Aug 4
This is a fully remote position, open to applicants in Latvia.
• Lead, mentor, and develop a team of 3 engineers dedicated to AI platform and product engineering.
• Drive the architectural vision for the AI product, prioritizing inference performance, reliability, cost efficiency, and developer experience.
• Actively contribute to the design, implementation, and optimization of model serving, inference pipelines, retrieval, orchestration, and evaluation systems.
• Take ownership of the AI/ML platform from start to finish, encompassing GPU capacity and scheduling, model deployment and versioning, data and embedding pipelines, monitoring, and quality assessment.
• Collaborate with Product, Infrastructure, Security, and Backend teams to establish priorities and facilitate seamless integrations.
• Conduct regular one-on-one meetings, performance evaluations, and technical coaching sessions.
• Promote ownership, engineering excellence, and a culture of continuous improvement.
• Convert strategic goals into actionable execution plans while ensuring timely delivery.
• Impact the evolution of the AI platform through pragmatic technical choices.
• Advocate for best practices in software development, automation, evaluation, and operational excellence.
• Over 7 years of experience in software engineering.
• At least 2 years in a technical leadership or engineering management position.
• Demonstrated experience in building and managing AI/ML platforms or inference infrastructure in a production environment.
• Proficiency in model serving, GPU workloads, latency optimization, and cost management.
• Experience in constructing distributed backend systems.
• Strong skills in Python.
• Familiarity with systems languages such as Go or Rust is advantageous.
• Knowledge of model APIs, open-weight models, vLLM, TGI, Triton, retrieval/vector databases, and prompt or context engineering.
• Understanding of evaluation frameworks, benchmarking, regression testing, and observability for non-deterministic systems.
• Extensive expertise in cloud platforms, preferably AWS, including EC2, ECS, Lambda, IAM, and CloudWatch.
• Experience with GPU computing.
• Familiarity with Kubernetes, containerization, and Infrastructure-as-Code tools like Terraform, Pulumi, or AWS CDK.
• Understanding of CI/CD automation tools such as GitHub Actions or CircleCI.
• Strong knowledge of Linux systems, networking basics, and monitoring/alerting best practices.
• Exceptional communication and leadership abilities.
• Strong technical judgment with the capacity to balance coding responsibilities with people management and strategic tasks.
• Proficiency in English is required.
• Fully remote work setup.
• Flexibility in a remote-first environment.
• Opportunity to engage with state-of-the-art LLMs, inference infrastructure, GPU computing, and distributed systems.
• Opportunities for leadership and technical advancement.
• Collaborative cross-functional work with Product, Infrastructure, Security, and Backend teams.
Remote People
Bestow
Virta Health
Space Inch
Get handpicked remote jobs straight to your inbox weekly.