Sr. Lead AI Engineer (Gen AI Platform Services)
Job Description
Capital One is hiring onsite in McLean, VA for a Sr Lead AI Engineer within Gen AI Platform Services. The role offers a competitive salary range of USD 229,900 - 262,400 per year, comprehensive health benefits, and a performance-based incentive program.
Responsibilities
- Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that change how our associates work and how our customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability, among others.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and introduce state of the art LLM optimization techniques to improve the performance metrics of large scale production AI systems, including scalability, cost, latency, and throughput.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Requirements
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies.
- At least 6 years of experience programming with Python, Go, Scala, or Java.
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- Nemo Guardrails
- PyTorch
- Python
- Go
- Scala
- Java
- Golang
Benefits
- Health benefits package
- Performance based incentive compensation, including cash bonuses and long term incentives
The ideal candidate
- You love to build systems, take pride in the quality of your work, and share our aim to help change banking for good.
- You stay current with the latest research and can translate scientific publications into production ready techniques.
- You adapt quickly, bring clarity to large, undefined problems, ask questions, and dig deep to uncover root causes, communicating findings concisely. You are willing to share new ideas even when unproven.
- You are deeply technical with a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI helps you spot optimization opportunities others miss.
- You are a resilient trail blazer who can forge new paths to achieve business goals when the route is unknown.
Preferred qualifications
- 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (AWS, Google Cloud, Azure, or equivalent private cloud)
- Experience designing, developing, integrating, delivering, and supporting complex AI systems
- Ability to lead and mentor an engineering team and influence cross functional stakeholders
- Experience developing AI and ML algorithms or technologies (LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang
- Experience applying state-of-the-art techniques to optimize training and inference to improve hardware utilization, latency, throughput, and cost
- Ongoing enthusiasm for AI research and systems, with the ability to judiciously apply novel techniques in production
- Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers