This position is no longer accepting applications
Closed on August 15, 2026.
This role is filled — get an email when new Engineering roles open on EngineerJobs.io:
Lead AI Engineer (GenAI Platform, AI Foundations, LLM Core and Agentic AI)
Agentic Ai
Ai Ml
Ai Platform
Artificial Intelligence
Engineering
Foundation Models
Genai
Large Language Models
Machine Learning
Scala
Software Engineering
Technical Lead
View similar jobs
Get alerted when similar jobs are posted — set up a New Engineering jobs on EngineerJobs.io alert.
See other roles at Capital One.
Job Description
Capital One seeks a Lead AI Engineer focused on GenAI Platform, AI Foundations, LLM Core and Agentic AI, onsite in Richmond, VA.
Responsibilities
- Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that transform how associates work and how customers engage with Capital One.
- Design, build, validate, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and apply state of the art LLM optimization techniques to improve scalability, cost, latency, and throughput for large scale production AI systems.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Requirements
- Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies.
- At least 4 years of experience programming with Python, Go, Scala, or Java.
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- Nemo Guardrails
- PyTorch
- Python
- Go
- Scala
- Java
Benefits
- Health benefits
- Performance-based incentive compensation (cash bonuses and long-term incentives)
The Ideal Candidate
- You enjoy building systems, take pride in the quality of your work, and share a commitment to advancing banking for good.
- You stay current with the latest research and can translate publications into production ready approaches.
- You adapt quickly, bring clarity to complex problems, ask questions, and articulate findings concisely; you welcome new ideas even if unproven.
- You are deeply technical with a strong foundation in engineering and mathematics, and you leverage hardware, software, and AI expertise to spot optimization opportunities others miss.
- You are a resilient trailblazer who forges new paths to achieve business goals when the route is uncertain.
Preferred Qualifications
- 6 or more years deploying scalable and responsible AI solutions on cloud platforms (AWS, Google Cloud, Azure, or equivalent private cloud).
- Experience designing, developing, delivering, and supporting AI services.
- Experience developing AI and ML algorithms or technologies (LLM inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang.
- Experience developing and applying state-of-the-art techniques to optimize training and inference software for better hardware utilization, latency, throughput, and cost.
- Passion for staying current with AI research and AI systems, and applying novel techniques in production.