Senior Lead AI Engineer (Gen AI Platform Services)
Job Description
Capital One invites applications for a Senior Lead AI Engineer specializing in Gen AI Platform Services. This role focuses on designing and delivering AI powered software components, including foundation model training, LLM inference, guardrails, governance, and observability, while guiding the long term roadmap for foundational AI systems.
Job Details
Location: New York, NY (onsite)
Salary: USD 250,800 - 286,200 per year
Experience: 4+ years required
Education: Master’s degree required
Responsibilities
- Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI driven products that transform internal workflows and customer interactions at Capital One.
- Design, build, test, deploy, and maintain AI software components such as foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Utilize a broad stack of open source and SaaS AI technologies including AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Develop and apply state of the art LLM optimization techniques to improve performance metrics such as scalability, cost efficiency, latency, and throughput in production AI systems at scale.
- Contribute to the technical vision and the long term roadmap for foundational AI systems at Capital One.
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- Nemo Guardrails
- PyTorch
- Python
- Go
- Scala
- Java
Overview
Capital One is focused on building responsible and reliable AI systems that advance banking for good. The company has established leadership in applying machine learning to deliver real time, personalized customer experiences. Ongoing investments in technology infrastructure and top talent position Capital One at the forefront of AI driven enterprise solutions, enabling capabilities across customer service and product innovation.
Team Description
The Intelligent Foundations and Experiences (IFX) team acts at the core of Capital One's AI initiatives. The group collaborates with partners across the enterprise to advance AI research and engineering, developing proprietary solutions that scale to millions of customers. AI models and platforms built by the team empower other teams to enhance products through AI.
In This Role, You Will
- Partner with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI driven products that transform how associates work and how customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and introduce advanced LLM optimization techniques to enhance scalability, cost efficiency, latency, and throughput for large scale production AI systems.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
The Ideal Candidate
- Enjoy building systems with pride in quality and align with the goal of improving banking outcomes.
- Stay current with the latest AI research and tactically apply novel techniques in production.
- Adapt quickly, bring clarity to complex problems, and communicate findings concisely; willing to propose new ideas even if unproven.
- Possess a deep technical foundation in engineering and mathematics, with the ability to recognize optimization opportunities across hardware, software, and AI.
- Act as a trailblazer who can forge paths to achieve business goals in uncertain environments.
Basic Qualifications
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related field plus at least 6 years of AI/ML development experience, or a Master’s degree in the same fields plus at least 4 years of AI/ML development experience.
- At least 6 years of programming experience in Python, Go, Scala, or Java.
Preferred Qualifications
- Seven years of experience deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, Azure, or equivalent private clouds.
- Experience designing, developing, integrating, delivering, and supporting complex AI systems.
- Proven ability to lead and mentor an engineering team and influence cross functional stakeholders.
- Experience developing AI and ML algorithms or technologies (for example LLM inference, similarity search and VectorDBs, guardrails, memory) using Python, C++, C#, Java, or Golang.
- Experience optimizing training and inference software to improve hardware utilization, latency, throughput, and cost.
- Strong interest in AI research and systems, with ability to judiciously apply new techniques in production environments.
- Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers.