Description
Capital One is hiring a Sr. Lead AI Engineer focused on inference optimization, foundation model hosting, and AI platform development. The role partners with engineers, scientists, program managers, and product managers to design, develop, test, deploy, and support AI software components, including large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability. The engineer will use technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, and PyTorch, and will develop techniques to improve the scalability, cost, latency, and throughput of production AI systems. The position requires a relevant bachelor's or master's degree plus the specified years of AI/ML and programming experience, and is available in Cambridge, McLean, New York, San Francisco, and San Jose, with stated annual salary ranges and visa sponsorship consideration.
