
Staff AI Engineer
Posted 4 days ago

Posted 4 days ago
This is a fully remote position, open to applicants in California, +3 more states.
• Collaborate with engineers, research scientists, technical program managers, and product managers to create AI-driven products.
• Design, develop, test, deploy, and support AI software components, including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
• Utilize Open Source and SaaS AI technologies such as AWS Ultraclusters, Hugging Face, VectorDBs, and PyTorch.
• Innovate and implement foundation model optimization techniques to enhance scalability, cost-efficiency, latency, and throughput of large-scale production AI systems.
• Contribute to the technical vision and long-term strategy of foundational AI systems.
• Define the technical direction for organization-wide AI architecture, tooling, observability, and deployment standards.
• Take ownership of the design and integration of model routing, caching, and orchestration systems for hybrid and multi-model workloads.
• Advocate for responsible AI principles, including transparency, reproducibility, and fairness-by-design.
• Promote internal education, mentorship, and the dissemination of best practices through architecture councils and AI guilds.
• Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields with a minimum of 8 years of experience in developing AI and ML algorithms or technologies; or a Master’s degree in these fields with at least 6 years of experience.
• Minimum of 8 years of programming experience in Python, Go, Scala, CUDA, or Java.
• Demonstrated experience in designing AI systems with considerations for cost, latency, throughput, and accuracy trade-offs.
• Experience in deploying scalable and responsible AI solutions on cloud platforms, including AWS, Google Cloud, Azure, or a comparable private cloud.
• Proven experience in architecting, designing, developing, integrating, delivering, and supporting complex AI systems.
• Experience in leading and mentoring multiple engineering teams and influencing cross-functional stakeholders up to the VP level.
• Proficiency in developing AI and ML algorithms or technologies using Python, C++, C#, Java, CUDA, or Golang.
• Familiarity with developing and implementing state-of-the-art techniques for optimizing training and inference software.
• Experience in building agentic AI systems and workflows.
• Knowledge of defining and operationalizing enterprise AI architecture standards, including data pipeline governance, observability, and evaluation frameworks.
• Experience in leading federated or multi-cloud AI strategies.
• Proven ability to influence research-to-production promotion processes.
• Experience in defining north-star metrics for AI systems.
• Capable of right-sizing models, instance counts, and hardware types.
• Capital One is open to sponsoring a new qualified applicant for employment authorization.
• Performance-based incentive compensation, which may include cash bonuses and/or long-term incentives (LTI).
• A comprehensive, competitive, and inclusive package of health, financial, and other benefits that support total well-being.
• Reasonable accommodations available for applicants who need them.
Alzheimer's Association®
Capital One
Capital One
Get handpicked remote jobs straight to your inbox weekly.