
Staff AI Engineer
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in California, +3 more states.
• Collaborate with engineers, research scientists, technical program managers, and product managers to create AI-driven products.
• Design, develop, test, deploy, and maintain AI software components, including foundation model training, LLM inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
• Utilize open-source and SaaS AI technologies such as AWS Ultraclusters, Hugging Face, VectorDBs, and PyTorch.
• Innovate and implement cutting-edge foundation model optimization techniques.
• Enhance scalability, cost efficiency, latency, and throughput of large-scale production AI systems.
• Contribute to the technical vision and the long-term roadmap for foundational AI systems.
• Establish the technical direction for enterprise-wide AI architecture and standardize tooling, observability, and deployment practices.
• Take ownership of the design and integration of model routing, caching, and orchestration systems for hybrid and multi-model workloads.
• Advocate for responsible AI principles, ensuring transparency, reproducibility, and fairness-by-design.
• Promote internal education, mentorship, and the sharing of best practices through architecture councils and AI guilds.
• Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field, with a minimum of 8 years of experience in developing AI and ML algorithms or technologies; or a Master's degree in one of these areas plus at least 6 years of experience.
• Minimum of 8 years of programming experience in Python, Go, Scala, CUDA, or Java.
• Proven experience in designing AI systems with considerations for cost, latency, throughput, and accuracy trade-offs.
• Familiarity with deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, Azure, or similar private clouds.
• Experience in architecting, designing, developing, integrating, delivering, and supporting intricate AI systems.
• Demonstrated leadership and mentorship skills across multiple engineering teams, influencing cross-functional stakeholders up to the VP level.
• Background in developing AI and ML algorithms or technologies utilizing Python, C++, C#, Java, CUDA, or Golang.
• Expertise in optimizing training and inference software for hardware utilization, latency, throughput, and cost efficiency.
• Experience in building agentic AI systems and workflows.
• Knowledge of AI research and the application of innovative techniques in production environments.
• Exceptional communication and presentation skills for conveying complex AI concepts effectively.
• Proven track record in defining and implementing enterprise AI architecture standards, including data pipeline governance, observability, and evaluation frameworks.
• Experience in leading federated or multi-cloud AI strategies.
• Ability to influence the processes of promoting research to production.
• Capability in defining north-star metrics for AI systems.
• Experience in right-sizing models, instance counts, and hardware types.
• Capital One is open to sponsoring employment authorization for a new qualified applicant.
• Performance-based incentive compensation, which may encompass cash bonuses and/or long-term incentives (LTI).
• Comprehensive health, financial, and other benefits that support overall well-being.
• Employment authorization sponsorship may be available for a new qualified applicant.
• Reasonable accommodations will be provided for applicants who require them.
Gainwell Technologies
SmartLogic
Scale Army Careers
Get handpicked remote jobs straight to your inbox weekly.