EngineerJobs.io
← Back to all jobs

Job Description

Capital One seeks a Lead AI Engineer focused on GenAI Platform, AI Foundations, LLM Core and Agentic AI, onsite in Richmond, VA.

Responsibilities

  • Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that transform how associates work and how customers engage with Capital One.
  • Design, build, validate, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
  • Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
  • Invent and apply state of the art LLM optimization techniques to improve scalability, cost, latency, and throughput for large scale production AI systems.
  • Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.

Requirements

  • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies.
  • At least 4 years of experience programming with Python, Go, Scala, or Java.

Technologies

  • AWS Ultraclusters
  • Huggingface
  • VectorDBs
  • Nemo Guardrails
  • PyTorch
  • Python
  • Go
  • Scala
  • Java

Benefits

  • Health benefits
  • Performance-based incentive compensation (cash bonuses and long-term incentives)

The Ideal Candidate

  • You enjoy building systems, take pride in the quality of your work, and share a commitment to advancing banking for good.
  • You stay current with the latest research and can translate publications into production ready approaches.
  • You adapt quickly, bring clarity to complex problems, ask questions, and articulate findings concisely; you welcome new ideas even if unproven.
  • You are deeply technical with a strong foundation in engineering and mathematics, and you leverage hardware, software, and AI expertise to spot optimization opportunities others miss.
  • You are a resilient trailblazer who forges new paths to achieve business goals when the route is uncertain.

Preferred Qualifications

  • 6 or more years deploying scalable and responsible AI solutions on cloud platforms (AWS, Google Cloud, Azure, or equivalent private cloud).
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience developing AI and ML algorithms or technologies (LLM inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang.
  • Experience developing and applying state-of-the-art techniques to optimize training and inference software for better hardware utilization, latency, throughput, and cost.
  • Passion for staying current with AI research and AI systems, and applying novel techniques in production.

Similar Jobs