
Staff AI Engineer
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in California, +2 more states.
• Spearhead the architecture for production AI systems.
• Create pipelines for model serving, inference, evaluation, and optimization.
• Enhance model latency, throughput, quality, and cost efficiency.
• Drive technical AI projects across various teams.
• Set engineering standards for the development and deployment of AI systems.
• Guide and mentor AI engineers.
• Collaborate with Product and Engineering leadership on AI strategies.
• Assess emerging models and infrastructure technologies.
• Minimum of 7 years of engineering or machine learning experience.
• Profound expertise in Python and ML engineering.
• Strong experience with production-level large language models (LLMs).
• Excellent skills in distributed systems and system design.
• Proven experience in optimizing inference systems.
• Strong technical leadership qualities and effective communication skills.
• Proficiency in English is required.
• Preferred: Familiarity with vLLM, TGI, Triton, CUDA, and PyTorch.
• Preferred: Experience with Kubernetes and GPU infrastructure.
• Preferred: Knowledge of RAG and fine-tuning.
• Preferred: Experience with AI agents.
• Preferred: Background in high-scale AI production environments.
• Comprehensive benefits package.
• Competitive salary.
• A collaborative and fast-paced work environment.
• Significant impact on architecture, product direction, and engineering culture.
• Opportunity for growth alongside a driven team.
Get handpicked remote jobs straight to your inbox weekly.