EngineerJobs.io
← Back to all jobs

Job Description

Lead AI Engineer for FM Hosting and LLM Inference at Capital One, onsite in McLean, VA, focused on building responsible AI systems and delivering AI powered products at scale.

Responsibilities

  • Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that transform how our associates work and how our customers interact with Capital One.
  • Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
  • Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
  • Invent and apply state of the art LLM optimization techniques to improve performance metrics like scalability, cost, latency, and throughput in production AI systems.
  • Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.

Basic Qualifications

  • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related field plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related field plus at least 2 years of experience developing AI and ML algorithms or technologies.
  • At least 4 years of experience programming with Python, Go, Scala, or Java.

Preferred Qualifications

  • 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (AWS, Google Cloud, Azure, or equivalent private cloud).
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang.
  • Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost.
  • Passion for staying abreast of the latest AI research and AI systems, and judiciously applying novel techniques in production.

Technologies

  • AWS Ultraclusters
  • Huggingface
  • VectorDBs
  • Nemo Guardrails
  • PyTorch
  • Python
  • Go
  • Scala
  • Java

Benefits

  • Health benefits
  • Financial benefits
  • Performance based incentive compensation (cash bonuses and/or long term incentives)

Team Description

The Intelligent Foundations and Experiences (IFX) team sits at the core of Capital One's AI strategy, partnering across the company to advance state of the art in AI engineering and to build and deploy proprietary solutions that deliver value to millions of customers. Our AI models and platforms empower teams to enhance products with AI capabilities at scale.

The Ideal Candidate

  • You love to build systems, take pride in the quality of your work, and are motivated to do the right thing. You want to tackle problems that can positively change banking.
  • Strong appetite for the latest AI research and the ability to interpret scientific publications to judiciously apply novel techniques in production.
  • You adapt quickly, bring clarity to large, undefined problems, ask questions, and articulate findings concisely; you share new ideas even when they are unproven.
  • Deep technical foundation in engineering and mathematics with expertise across hardware, software, and AI to spot optimization opportunities others miss.
  • A resilient trailblazer who can forge new paths to achieve business goals when the route is unknown.

Similar Jobs