Lead AI Engineer (GenAI Platform, AI Foundations, LLM Core and Agentic AI)
Job Description
Capital One seeks a Lead AI Engineer focused on GenAI Platform, AI Foundations, LLM Core and Agentic AI, onsite in Richmond, VA.
Responsibilities
- Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that transform how associates work and how customers engage with Capital One.
- Design, build, validate, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and apply state of the art LLM optimization techniques to improve scalability, cost, latency, and throughput for large scale production AI systems.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Requirements
- Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies.
- At least 4 years of experience programming with Python, Go, Scala, or Java.
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- Nemo Guardrails
- PyTorch
- Python
- Go
- Scala
- Java
Benefits
- Health benefits
- Performance-based incentive compensation (cash bonuses and long-term incentives)
The Ideal Candidate
- You enjoy building systems, take pride in the quality of your work, and share a commitment to advancing banking for good.
- You stay current with the latest research and can translate publications into production ready approaches.
- You adapt quickly, bring clarity to complex problems, ask questions, and articulate findings concisely; you welcome new ideas even if unproven.
- You are deeply technical with a strong foundation in engineering and mathematics, and you leverage hardware, software, and AI expertise to spot optimization opportunities others miss.
- You are a resilient trailblazer who forges new paths to achieve business goals when the route is uncertain.
Preferred Qualifications
- 6 or more years deploying scalable and responsible AI solutions on cloud platforms (AWS, Google Cloud, Azure, or equivalent private cloud).
- Experience designing, developing, delivering, and supporting AI services.
- Experience developing AI and ML algorithms or technologies (LLM inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang.
- Experience developing and applying state-of-the-art techniques to optimize training and inference software for better hardware utilization, latency, throughput, and cost.
- Passion for staying current with AI research and AI systems, and applying novel techniques in production.